Gemini 3.8 Flash posted scoring 73.7% on DeepSWE 1.1
doodlestein · x · 2026-09-03
Google's Logan Kilpatrick shared that Gemini 3.8 Flash scores 73.7% on the DeepSWE 1.1 coding benchmark. Developer @doodlestein noted that 3.7 Flash was already respectable and expressed excitement to try 3.8. This is an early score share with no full evaluation report yet.
Related event: Gemini 3.8 Flash scores 73.7% on DeepSWE 1.1 benchmark(5 posts)→
More from Models
- ChatGPT helps resolve 30-year-old stable forking conjecture in logic — Dr_Singularity · 2026-09-03
- Users praise Gemini 3.8 Flash: "coding is fun again" — xennygrimmato_ · 2026-09-03
- Grok Bot's Token Limits Frustrate Users; Subscriptions With All-You-Can-Eat Tiers May Win — NickPassig · 2026-09-03
- Author Burkov says Anthropic ignored his refund requests, so he cobbles together Codex and Grok — burkov · 2026-09-03
- Google launches Gemini 3.8 Flash Cyber security model alongside Fairwind Program for defenders — GoogleAI · 2026-09-03
- Matt Shumer: slow AI releases aren't a wall — safety clearance is the bottleneck, wave of frontier models imminent — mattshumer_ · 2026-09-03