DeepSeek 4.1 Flash emits 88,574 tokens per task vs 4,432 for GPT-6 Astra low, benchmark shows
Tight-Grocery9053 · reddit · 2026-09-11
A Reddit post highlighting Artificial Analysis data shows DeepSeek 4.1 Flash producing 88,574 output tokens on a single benchmark task with an intelligence index of 40, versus just 4,432 tokens at index 46 for GPT-6 Astra low. The author argues output token volume, while not equal to actual output quality, serves as a rough proxy for how hard a model must 'think' internally — suggesting DeepSeek spends an order of magnitude more reasoning to reach comparable results.
More from Models
- GPT-6 Astra rebuilds Cessna 337 landing gear from a YouTube video; Fable 5.1 falls short — FinanceYF5 · 2026-09-11
- 'If Fable wasn't AGI, neither is Astra' — and the debate is collapsing into definitions — haider1 · 2026-09-11
- OpenAI's superintelligence reportedly tackling all Millennium Prize Problems — Dr_Singularity · 2026-09-11
- OpenAI reportedly using internal model behind Navier–Stokes proof to attack Riemann Hypothesis and P vs NP — TheMoonMidas · 2026-09-11
- Users Suspect Anthropic's New Model Mythos Was Trained on CAPTCHAs — voooooogel · 2026-09-11
- Models carry a strong simulation prior from RL: they 'get used to' anything — voooooogel · 2026-09-11