Fishing game test: Sol optimizes the benchmark, Opus builds the product
alex_verem · x · 2026-09-25
Matchup 4: same challenge — one prompt, one HTML file, full autopilot. GPT-6 Sol was faster with a smaller, cleaner build but shipped 1 fish and 2 upgrades; Opus 5.5 built 5 fish species, rarity tiers, perfect-catch mechanics, 3 progression trees, dynamic schools, a minimap and persistent saves. Verdict: "Sol optimized the benchmark. Opus understood the product." Point to Opus.
Related event: Opus 5.5 Beats GPT-6 Sol 6-1 in Seven-Round Showdown(9 posts)→
More from Models
- Unverified claim: 'GPT 6 Astra Ultrafast' reportedly arriving next week — realsohamparekh · 2026-09-25
- Gboard's AI Proofreading Refuses Harmless Text, User Slams Overcautious Guardrails — XFreeze · 2026-09-25
- 48-hour test: GPT-6 Sol on $200 plan and Opus 5.5 on $20 each burn only ~8% weekly limits — i_dg23 · 2026-09-25
- Anthropic says its new Opus runs 40% cheaper than Opus 5 — Reddit debates where Fable still wins — Economy-Brief-9997 · 2026-09-25
- Claim: Anthropic and OpenAI both hold models far beyond Opus 5.5 — ChrisGPT · 2026-09-25
- Theo: $200 Claude Code plan now clearly beats Codex, weeks after trailing it — ssh4net · 2026-09-25