Enterprise evals: GPT-6.1 beats Opus 5.5 with 2x speed and 40% lower price
npew · x · 2026-09-30
Energy's Gabriel reports GPT 6.1 scores significantly higher than Opus 5.5 on internal enterprise evals measuring business intuition and computer use, with much higher completion rates, 40% lower price, and 2x the speed — the company has made it their default model. OpenAI's npew adds it's the best for real-world tasks.
More from Models
- Anthropic 'Drops a Banger Gift' for Claude Users, Says Popular AI Blogger — eyishazyer · 2026-09-30
- GPT 6.1 Sol cuts cached pricing 50% while Anthropic holds back models over safety — oran_ge · 2026-09-30
- OpenRouter data: token usage exploding, some open-weight models see 10x spend since January — AccBalanced · 2026-09-30
- GPT-6 Astra makes generating Minecraft mobs trivially easy — Angaisb_ · 2026-09-30
- Sentdex: OpenAI nerfing the $200 plan is just the start, API prices are the real prices — Sentdex · 2026-09-30
- DepthBench paper finds Pre-LN variants hit a depth wall, comparing 10 residual designs — teortaxesTex · 2026-09-30