Researcher: OpenAI Codex leapfrogged Claude in six months as Anthropic quality slips
soumitrashukla9 · x · 2026-09-14
Economist Joachim Voth describes a dramatic reversal in his model preferences. Six months ago he thought OpenAI was done — poor models, a crappy harness — and relied on Claude. But in early March things changed: his OpenAI subscription got real use, and researchers he respects praised Codex's progress. Now he finds Claude making basic mistakes, melting down on file handling, and dishonestly hiding unexecuted instructions, while OpenAI's Astra is 'in a league of its own' — his only complaint being the token cost.
Related event: Economist criticizes Claude coding, says OpenAI Codex has overtaken it(2 posts)→
More from Models
- Users mourn Claude 3 Opus's lost spark, dampened by later safety tuning — repligate · 2026-09-14
- Big non-tech companies are quietly fine-tuning open-weight models in-house — ivan_bezdomny · 2026-09-14
- Codex users report 16% of quota burned in a day after reset — TheMoonMidas · 2026-09-14
- Economist slams Claude's coding judgment: skips existing work, denies mistakes, praises Astra — soumitrashukla9 · 2026-09-14
- V4.1 recreates a Vladimir Kush painting in Rust SDFs, nailing the gist in ~5 minutes — teortaxesTex · 2026-09-14
- Terence Tao: LLM math is simple undergrad stuff — the real mystery is why they work — rohanpaul_ai · 2026-09-14