Gemini web UI swaps thinking toggle for tiered effort levels, seen as prep for Gemini 4 Argon
PeteyPab305 · reddit · 2026-10-10
Gemini's web UI has replaced the binary Extended Thinking toggle with granular Low/Medium/High effort levels under the 3.5 Flash-Lite, 3.8 Flash, and 3.1 Pro pickers. Low strips deliberation overhead for instant streaming; High allocates maximum test-time compute for complex debugging and multi-step logic. The author argues this gives direct control over latency and token burn and is groundwork for the rumored Gemini 4 Argon, which is built around long-horizon execution and sustained chains of thought—something a rigid on/off switch couldn't support without blowing quotas or tanking response times.
Related event: Gemini Replaces Thinking Toggle with Three Effort Levels(4 posts)→
More from Models
- Claude iOS app's unkillable pseudo-notifications draw fire from ex-OpenAI safety VP — Miles_Brundage · 2026-10-11
- Google's EmbeddingGemma 2 runs free and fully local on a Mac — Saboo_Shubham_ · 2026-10-11
- User says Google AI Mode still hallucinates after a year without seeing AI errors — burny_tech · 2026-10-11
- User burneda $17 Claude bill before a task even finished — lxfater · 2026-10-11
- Self-funded AI user: cost per task is the only benchmark that matters — victor_explore · 2026-10-11
- Mathematicians reviewing OpenAI's Navier-Stokes proof find the Lean formalization sound — burny_tech · 2026-10-11