Gemini 3.8 Flash outputs 48k tokens per task on high reasoning, 30% more than predecessor
ArtificialAnlys · x · 2026-09-02
Artificial Analysis notes Gemini 3.8 Flash averages 48k output tokens per task on high reasoning, up 30% from 3.7 Flash, at 300 tok/s; Time per Task rises from 2.2 to 2.5 minutes, behind Claude Fable 5.1 (2.1 min). Low reasoning outputs 14k tokens per task in 0.8 minutes.
Related event: Gemini 3.8 Flash Benchmarks: Smarter, Pricier, Still on the Cost Frontier(8 posts)→
More from Models
- GPT-6-Astra model slug spotted on OpenAI APIs, hinting at routing tests — testingcatalog · 2026-09-03
- Polymarket puts 78% odds on OpenAI's rumored Astra model launching tomorrow — Polymarket · 2026-09-03
- Polymarket puts 78% odds on OpenAI's rumored Astra model shipping tomorrow — Polymarket · 2026-09-03
- Unverified report: Gemini 3.8 Flash out with big agentic coding gains — clmt · 2026-09-03
- 'GPT-6-ASTRA' spotted staged on the OpenAI API, unconfirmed — ThunderBeanage · 2026-09-03
- Google DeepMind releases Gemini 3.8 Flash and 3.8 Flash Cyber — Google DeepMind · 2026-09-03