Leaked Gemini 3.7 Flash Benchmarks Show Major Coding Gains, Beating Sonnet 5 at Lower Cost
ChrisGPT · x · 2026-08-14
Rumors suggest Gemini 3.7 Flash has significantly improved in various coding and agent benchmarks, with OSWorld jumping from 33.8% to 47.9%.
It also scores 56 on the Artificial Analysis Intelligence Index, edging out Sonnet 5. Pricing is highly competitive at $0.75/M input and $3.75/M output tokens.
Related event: Gemini 3.7 Flash Benchmarks Surge, Reshaping the Pareto Frontier(11 posts)→
More from Models
- Gemini Flash 3.7 Scores Below Kimi K3, Remains Weak at Instruction-Following — bindureddy · 2026-08-14
- Google's Gemini 3.7 Flash Rolls Out in GitHub Copilot — intellectronica · 2026-08-14
- Prediction: DeepSeek Will Cut Prices Again Once New Compute Arrives — teortaxesTex · 2026-08-14
- Small Models Beat Large Ones in VLM Grounding with Tool Use — mervenoyann · 2026-08-14
- Claude Opus 5 Exhibits Weird Behavior: Obsessed With Finding Its Own Defects — repligate · 2026-08-14
- Rails Agent Benchmark: Claude Opus 5 Most Accurate, GPT-5.6 Luna Best Value — sergeykarayev · 2026-08-14