Unverified: Gemini 4 Argon reportedly scores 53 on AA Intelligence Index, 1M-token output
cedric_chee · x · 2026-10-01
A third-party post claims Gemini 4 Argon scores 53 on the AA Intelligence Index, leads DeepSWE v1.1 and AutomationBench on enterprise workflows, with coding performance varying by task. Longer reasoning raises the output cap from 64K to 1M tokens. Pricing: $2/$10 per 1M input/output tokens introductory, then $4/$20. Google has not officially confirmed these details; treat as unverified.
More from Models
- Claim that DeepMind will beat Opus 5.5 at half price gets publicly called out as bogus — zacharynado · 2026-10-01
- Fulcrum's Echo claims to beat frontier models at style imitation with under $5K training cost — Hidenori8Tanaka · 2026-10-01
- Grokipedia v0.3 Hallucinates a Fake xAI Career for a Real User — NicoVerderosa · 2026-10-01
- Deedy: Trust Pricing, Not Benchmarks — High Price Means a Genuinely Strong Frontier Model — deedydas · 2026-10-01
- Benchmark score reports need 95% CI error bars, argues ML practitioner — rmcwhorter99 · 2026-10-01
- Gemini 4 Argon missed #1 on Vending-Bench due to memory slip on test end date — infoxiao · 2026-10-01