DeepSeek 4.1 Flash emits 88,574 tokens per task vs 4,432 for GPT-6 Astra low, benchmark shows

Tight-Grocery9053 · reddit · 2026-09-11

A Reddit post highlighting Artificial Analysis data shows DeepSeek 4.1 Flash producing 88,574 output tokens on a single benchmark task with an intelligence index of 40, versus just 4,432 tokens at index 46 for GPT-6 Astra low. The author argues output token volume, while not equal to actual output quality, serves as a rough proxy for how hard a model must 'think' internally — suggesting DeepSeek spends an order of magnitude more reasoning to reach comparable results.

Original post →

More from Models

Models channel →