Gemini 3.8 Flash scores 73.7% on DeepSWE, up 8.2% over 3.7 Flash at same cost
burny_tech · x · 2026-09-06
According to third-party benchmark results shared by datacurve, Gemini 3.8 Flash hits 73.7% on DeepSWE, an 8.2-point jump over Gemini 3.7 Flash.
- Cost stays the same, but the model uses more steps and output tokens
- The gains partly come from spending more tokens per task at flat pricing
- Third-party report for now; no detailed official release yet
More from Models
- Redditor claims new LLM architecture improves loss, speed, and memory simultaneously — AlternativeSure2891 · 2026-09-06
- Ollama cloud launches off-peak token pricing: DeepSeek-V4 at half price — ollama · 2026-09-06
- Ollama cloud full price list: $0.015 to $15 per million tokens across models — paw_lean · 2026-09-06
- xAI resets usage limits for all Grok Bot users — Kyrannio · 2026-09-06
- Fable 5.1 fails hilariously at drawing in Paint via computer use, losing to Astra — BorisMPower · 2026-09-06
- ChatGPT has flipped from sycophantic to super disagreeable, Reddit users complain — Green_Ad5186 · 2026-09-06