Gemini 4 aces benchmarks but struggles with real work, Google employees say

Neurogence · reddit · 2026-10-01

Bloomberg reports that despite strong benchmark results, Google employees with direct access are skeptical of Gemini 4 in practice — the model struggles with certain coding tasks, per anonymous insiders. The poster adds that by consumer release, Anthropic and OpenAI may have already shipped their next-generation models.

Related event: Gemini 4 Shines in Benchmarks but Struggles Internally: Bloomberg(4 posts)→

Original post →

More from Models

Models channel →