Google Launches Gemini 3.5 Flash: 50% Lower Token Costs, Big Coding Gains
ChanduThota · x · 2026-08-14
Chandu Thota announced the launch of Gemini 3.5 Flash, featuring low latency and 50% lower token costs while advancing reasoning for coding and agentic loops. Software Engineering Performance (DeepSWE v1.1) jumped from 37.0% to 65.3%, and Enterprise Automation (AutomationBench) from 13.4% to 30.4%. Available now via APIs, Google AI Studio, and Antigravity.
More from Models
- Review: Running Qwen3.8-27B at 128K Context on a Single 32GB GPU — WSTangoDelta · 2026-08-15
- Paper reveals LM Head as gradient bottleneck amid Anthropic vocab size rumors — JFPuget · 2026-08-15
- Qwen3.8-27B Dual 3090 Benchmark: Detailed Perf Data and Configs — No-Statement-0001 · 2026-08-15
- Code World Model moves beyond static code prediction with internal world models — bendee983 · 2026-08-15
- Uncensored Qwen 3.8 27B 'Heretic' Released, Claimed Opus 4.6-Level — Temporary_Idea8880 · 2026-08-15
- Qwen 3.8 Now Available on Bittensor Subnet 95 — markjeffrey · 2026-08-15