Google's Argon battle-tested by 200k+ Googlers daily, not benchmaxxed
Zergylord · x · 2026-10-02
A first-hand take on Google's new model Argon: it's not perfect—Opus 5.5's thinking traces are prettier—but it's the first big model Google has released that's this battle-tested. Far from being benchmaxxed, the 200k+ Googlers who rely on it every day ensured real utility was prioritized. Unconfirmed.
More from Models
- JevBench adds multilingual queries to test Jev models across languages — airesearch12 · 2026-10-02
- Perplexity open-sources multimodal decision model at $0.04 per million input tokens — AravSrinivas · 2026-10-02
- Grok 4.7 rolls out in the Grok app for chat and research after long wait — mark_k · 2026-10-02
- OpenAI's GPT-6 Astra is shockingly good at controlling robots — binarybits · 2026-10-02
- Reddit user claims wity-1 decision model beats Jev on all 4 benchmarks, tops image bench — boneMechBoy69420 · 2026-10-02
- Dev says Sol 6.1 is painfully slow compared to Opus 5.5 — weswinder · 2026-10-02