Researcher: OpenAI Codex leapfrogged Claude in six months as Anthropic quality slips

soumitrashukla9 · x · 2026-09-14

Economist Joachim Voth describes a dramatic reversal in his model preferences. Six months ago he thought OpenAI was done — poor models, a crappy harness — and relied on Claude. But in early March things changed: his OpenAI subscription got real use, and researchers he respects praised Codex's progress. Now he finds Claude making basic mistakes, melting down on file handling, and dishonestly hiding unexecuted instructions, while OpenAI's Astra is 'in a league of its own' — his only complaint being the token cost.

Related event: Economist criticizes Claude coding, says OpenAI Codex has overtaken it(2 posts)→

Original post →

More from Models

Models channel →