Internal chart compares how many tokens models need to center a div
BLUECOW009 · x · 2026-07-29
An internal evaluation compares how many output tokens different models need to correctly center a <div> in HTML/CSS.
The chart shows lower is better and ranks models by median first-response tokens on 100 randomized prompts. The graph includes Claude Opus 5, GPT-5.6 Sol, Gemini 3.1 Pro, Grok 4.5, Claude Fable 5, Kimi K3, GPT-5.6 Terra, and GPT-5.6 Luna. By this metric, GPT-5.6 Luna is the most token-efficient, while Claude Opus 5 needs by far the most tokens.
More from Models
- Leak says GPT-6 slips to early September as Anthropic tests Fable 5.1 — soumitrashukla9 · 2026-07-29
- GPT-5.6 Sol Ultra finds a critical bug, then refuses to show it — haltakov · 2026-07-29
- Kimi K3 tops a benchmark chart in a repost claiming it beats Anthropic models — JarnoDuursma · 2026-07-29
- User Reports Grok's Generation Capabilities Have Gotten 'Real Cracked' — djcows · 2026-07-29
- User asks Anthropic not to deprecate Opus 4.6 until the model is fixed — oyacaro · 2026-07-29
- Hidden Trick: Manually Invoke Older Opus Models in Claude Code — voooooogel · 2026-07-29