Gemini 3.8 Flash hits 305 tokens/sec, nearly 2x the runner-up

NewVeterinarian5384 · reddit · 2026-09-03

Per Google's official blog, Gemini 3.8 Flash outputs 305 tokens per second — nearly double second-place Muse Spark 1.2 (154 t/s) and well ahead of GPT-5.6 Luna (126 t/s) — with an intelligence score of 59, just below the 60-66 leader group. If the speed holds in real API use, long coding-agent runs could feel far less painful. A 3.8 Flash Cyber variant shipped alongside.

Original post →

More from Models

Models channel →