Sonnet 5.5 vs GPT 6.1 Sol chess match via MCP: 18x more reasoning tokens, $16 vs $2.37

adigrazia80 · reddit · 2026-10-04

A Reddit user wired Sonnet 5.5 (Claude Desktop) and GPT 6.1 Sol (Codex) into a local chess app via MCP, having each play a full game in a single conversation with no engine or legal-move hints. Score is 1-1 after two games.

Key numbers (both at Medium reasoning):

Subscription usage also diverged sharply: Claude's $20/month five-hour budget went from 23% to 56% during the match, while Codex's weekly display on the $200/month plan stayed at 5%. Sonnet called its pawn promotion "unstoppable," later admitted missing the Bf3 defense, and had one move rejected then self-corrected; Codex also misread a passed pawn once. The author cautions two games prove nothing about playing strength and plans a rematch with colors swapped.

Original post →

More from Fun

Fun channel →