Inquirer‘Very unhealthy’ air reported in parts of NCRESPNOgunbowale saves Wings' season, breaks playoff record with 45 in OT classicDaily MaverickParamount gets court green light on Warner Bros deal, names Mattel’s Kreiz co-CEOESPN DeportesEN VIVO: Padres tiene ventaja sobre Cubs, pero Chicago aún está a tiempo de reaccionarPunchMohbad: IGP approves family’s petition for fresh probeThe Jerusalem PostIsraeli tourists rescued amid renewed fighting in Tigray, Ethiopia, FM says한겨레[속보] 이 대통령, 하준경 정책실장·문신학 경제성장수석 임명SözcüAKP'li Zeybekçi'nin fonlara kaptırdığı para belli olduCNN TürkSON DAKİKA... SPK'dan kritik karar: 1 milyon TL'ye kadar ara ödeme yapılacakBBC NewsThe Papers: 'Israel's 9/11 stopped' and 'That was a great life'RapplerLIVE UPDATES: Impeachment trial of Vice President Sara DuterteThe Hollywood ReporterNicolas Cage Says the MCU “Is Becoming a Well-Funded and Glorified WWE”
The Daily Newsstand · Free, Always
Thursday, October 1, 2026

Google grapples with employee skepticism about new Gemini 4

Translate
While Gemini 4 has performed well on benchmarks widely used to gauge model efficacy, it does less well when employees actually put it to work, according to people with direct access to the effort.

While Gemini 4 has performed well on benchmarks widely used to gauge model efficacy, it does less well when employees actually put it to work, according to people with direct access to the effort. | BLOOMBERG

Bloomberg

Oct 1, 2026

Alphabet’s Google has begun rolling out Gemini 4 Argon, its long-awaited flagship artificial intelligence model, but the company is grappling with internal skepticism over how well it performs in key areas, such as coding.

Google unveiled the model Wednesday to a small group of trusted cybersecurity partners and said it would expand access after additional testing, starting with paid subscribers. The company said Gemini 4 posted leading scores on several benchmark tests, including beating OpenAI’s Astra model on one that measures security skills.

But some insiders say those metrics don’t tell the whole story. While Gemini 4 has performed well on benchmarks widely used to gauge model efficacy, it does less well when employees actually put it to work, according to people with direct access to the effort. The model struggles to handle certain coding tasks, said the people, who requested anonymity to discuss an internal matter.

In a time of both misinformation and too much information,
quality journalism is more crucial than ever.
By subscribing, you can help us get the story right.

View the original on The Japan Times →

KioskNews shows a cleaned-up reading view extracted from the publisher’s page — the original always lives on their site, not ours.