DeepMind Senior Staff Engineer Calls Bloomberg's Gemini 4 Employee Gripe Report "BS"
MrZiruiWang · x · 2026-10-02
Bloomberg reported that while Gemini 4 "has performed well on benchmarks widely used to gauge model efficacy, it does less well when employees actually put it to work." A Senior Staff Research Engineer at Google DeepMind publicly called the report "bs."
The poster adds that every model draws lovers and haters since nothing is perfect — what model hasn't "struggled to handle certain coding tasks"? The insider denial adds suspense ahead of the release: whether real-world usage matches the benchmark scores will only be settled once Gemini 4 ships.
More from Models
- Historian revamps 10-year-old Edo-era sankin kōtai animation with Claude, wowed by design upgrade — tkasasagi · 2026-10-02
- Musk: 'Super Intelligence' Now Aces Accounting Tests AI Failed 18 Months Ago — elonmusk · 2026-10-02
- Is anyone still using Mirostat? Do newer models still need adaptive sampling? — ParvusNumero · 2026-10-02
- Local LLM benchmark: fine-tuned 9B model cuts latency 4x on RTX 5070 at same quality — Storge2 · 2026-10-02
- Fable 5.5 release 'impending, as soon as next week', claims Bindu Reddy — bindureddy · 2026-10-02
- Open weights will be good enough: why frontier AI smarts won't matter for daily life — Dan_Jeffries1 · 2026-10-02