Gemini 3.6 Flash halves task time while 3.5 Flash-Lite gets faster but pricier
ArtificialAnlys · x · 2026-07-21
Artificial Analysis says Gemini 3.6 Flash and Gemini 3.5 Flash-Lite both cut average time per task to about half of their predecessors.
- 3.6 Flash: about 304 output tokens/sec, roughly 1.3 minutes per task, and a cost-per-task drop of around 18%.
- 3.5 Flash-Lite: about 350 output tokens/sec, roughly 0.6 minutes per task, but its cost per task more than doubles.
- The post attributes the speedup to better token efficiency and faster output generation, with pricing changes driving the cost differences.
More from Models
- Google launches three new Gemini models, including a cybersecurity system — Polymarket · 2026-07-22
- Google says information agents are coming to AI Pro and Ultra this summer — gaganghotra_ · 2026-07-22
- Poolside’s Laguna S 2.1 gets a two-week free run on Nous Portal — NousResearch · 2026-07-22
- Qwen3.8 Max Preview looks substantially better in a side-by-side test with Kimi K3 — curiousily_ · 2026-07-22
- Moonshot’s Kimi K3 reaches #5 on MathArena as the top open model — xeophon · 2026-07-22
- Google launches Gemini 3.5 Flash Cyber for CodeMender, with limited access for governments — GoogleAI · 2026-07-22