DeepMind Senior Staff Engineer Calls Bloomberg's Gemini 4 Employee Gripe Report "BS"

MrZiruiWang · x · 2026-10-02

Bloomberg reported that while Gemini 4 "has performed well on benchmarks widely used to gauge model efficacy, it does less well when employees actually put it to work." A Senior Staff Research Engineer at Google DeepMind publicly called the report "bs."

The poster adds that every model draws lovers and haters since nothing is perfect — what model hasn't "struggled to handle certain coding tasks"? The insider denial adds suspense ahead of the release: whether real-world usage matches the benchmark scores will only be settled once Gemini 4 ships.

Related event: Gemini 4 Leaks Flood Out: Strong Benchmarks, but Bloomberg Reports Real-World Coding Struggles(36 posts)→

Original post →

More from Models

Models channel →