Skeptic pushes back on 'Google is so back' hype, citing unreliable benchmarks
Angaisb_ · x · 2026-10-01
The author pushes back on the wave of "Google is so back" posts: the new model isn't even usable yet, and Google's model benchmarks have a track record of unreliability — pointing to the Gemini 3 Pro episode as a cautionary precedent.
Related event: Skeptics push back on 'Google is so back' hype over new model benchmarks(2 posts)→
More from Models
- Early Muse hands-on: friendly and compliant but forgets formatting and never double-checks — ivan_bezdomny · 2026-10-01
- Dev claims 'Sonnet 5.5' hits 120-130tps with 100% pass rate on his end-to-end app benchmark — julianharris · 2026-10-01
- GPT-6.1 Sol sees highest demand ever as OpenAI doubles serving speed — pvncher · 2026-10-01
- GPT-6 Astra cracks 217-year-old cipher to Napoleon's general in 6 hours — arthurcolle · 2026-10-01
- OpenAI's 20x-to-10x quota cut is a push from Astra to Sol, not a better deal — CtrlAltDwayne · 2026-10-01
- Gemini 4 Argon tops APEX-Agents at 82.2% Pass@1, first model to break 80%, but burns 2.6M tokens per run — xennygrimmato_ · 2026-10-01