Astra-max hits 95% on eyebench-v3 at half the cost of Sol-max, tokens ~3.8x fewer
adonis_singh · x · 2026-09-05
A blogger reports Astra-max scoring 95% on the vision benchmark eyebench-v3, at roughly half the cost of Sol-max while outputting 3.8x fewer tokens and nearly doubling its score. He adds that on an intelligence-per-token basis it beats everything else even at low reasoning settings, and says he'll move on to harder benchmarks instead of making a v4. Note: unofficial third-party evaluation.
Related event: Astra-max scores 95% on eyebench-v3 at half the cost of Sol-max(3 posts)→
More from Models
- No usage reset on GPT-6 Astra launch day, developer calls out OpenAI — tobowers · 2026-09-05
- Debate: Models Fuzzily Recall Concepts, Not Text — SAE Features vs Edit-Distance Memorization — voooooogel · 2026-09-05
- Blogger feeds GPT6 Astra a PPT template, gets 30 conference slides with the right avatar — vista8 · 2026-09-05
- Dozens of GPT-6 Astra prompts collected via GPT 6 pro web search — vista8 · 2026-09-05
- Tencent Hunyuan Hy4 preview: 770B total/49B active, 1M context, Apache 2.0, day-0 vLLM — aftahi_ai · 2026-09-05
- Astra Draws Stunning Illustrations Stroke by Stroke, Not Generated — Tolopono · 2026-09-05