GPT-6.1 Sol benchmarks across all effort levels land, making GPT-6 Astra hard to justify
PawelHuryn · x · 2026-10-01
Pawel Huryn published scores for all effort levels of GPT-6.1 Sol on his benchmark site (n=3 for max and xhigh, n=2 for others), with additional n=3 runs still in progress. Considering the cost, he sees no reason to use GPT-6 Astra instead.
More from Models
- Rox benchmarks: Jev reranking beats GPT-5 Mini — 20x faster, 10x cheaper, 12% more accurate — hardimanjames · 2026-10-01
- Nat Lambert: More Frontier Labs Like Google's Gemini 4 Benefit Consumers — natolambert · 2026-10-01
- Qwen3.8-Flash-Next Cut 44% via REAP Hits 70% on Terminal-Bench 2.1 — rmonsurate · 2026-10-01
- Google launches Gemini 4 Argon, a cybersecurity model that tops prompt injection benchmarks — ralucaadapopa · 2026-10-01
- Model wars: OpenAI went from best model in the world to arguably third place in a week — signulll · 2026-10-01
- Rumor: Gemini 4 spotted with SOTA knowledge scores, coding near Astra/Fable 5.1 level — haider1 · 2026-10-01