Gemini 4 Argon reportedly hallucinates 15% vs GPT-6's 51-54%: honesty over accuracy
eyishazyer · x · 2026-10-01
A viral post claims the most underrated Gemini 4 Argon stat is honesty: a 15% hallucination rate versus 51% for GPT-6 Astra and 54% for GPT-6.1 Sol. Accuracy is lower (50% vs 63%), but the model reportedly admits when it doesn't know instead of bluffing. Note: the model names and figures are unverified.
More from Models
- Gemini 4 Argon Posts Strong Result on AA Coding Agent Index via Antigravity — prateeky2806 · 2026-10-01
- Anthropic: GLM-5.3 builds Chrome exploits nearly as well as unreleased Claude Mythos, with lax safeguards — matthew_d_green · 2026-10-01
- Full Artificial Analysis Intelligence Index results for Solar Mini 4 — ArtificialAnlys · 2026-10-01
- Upstage's Solar Mini 4 scores 24 on AA Intelligence Index, but costs ~5x GPT-6 Luna per task — ArtificialAnlys · 2026-10-01
- Dev red-teaming GLM 5.3 finds bizarre traces, suspects OpenRouter routed to a 1-bit quant on someone's DGX Spark — voooooogel · 2026-10-01
- Google engineer teases Gemini 4's surprisingly strong long-context and long-sequence generation — RubenEVillegas · 2026-10-01