Unverified claim: Gemini 3.8 Flash tops DeepSWE v1.1 coding benchmark
vedantmisra · x · 2026-09-03
A viral post claims Gemini 3.8 Flash has topped the DeepSWE v1.1 coding benchmark, with a screenshot attached. This is a third-party claim with no official confirmation from Google, so treat it as unverified until corroborated.
Related event: Gemini 3.8 Flash leaks with 73.7% on DeepSWE 1.1, nearing flagship models(6 posts)→
More from Models
- Google Search Console social properties show bizarre zero-click queries, SEOs suspect AI fan-out — gaganghotra_ · 2026-09-03
- Gemini 3.8 Flash ties for top of DeepSWE at 74%, but burns 1.34x more output tokens — haider1 · 2026-09-03
- Gemini 3.8 Flash scores 69.9% on Cursor bench at $2.38 per task — _philschmid · 2026-09-03
- GPT-Astra rumored to be a GPT-4-scale leap as Altman teases 'significant step forwards' — DavidmComfort · 2026-09-03
- Gemini 3.8 Flash hits 305 tokens/sec, nearly 2x the runner-up — NewVeterinarian5384 · 2026-09-03
- Reddit user says ChatGPT auto-reload credit re-enables itself after being turned off — Callicojacks · 2026-09-03