OpenAI Lags Behind Claude in Writing Experience
RaphaelDabadie · x · 2026-07-13
The author believes Sol and Fable represent a new round of model capability improvements, but the instruction-following experience in 5.6 is disappointing for writing tasks. Meanwhile, they still feel Claude outperforms OpenAI in coherence and clarity, though both models tend to overuse unnecessary jargon. The author concludes by urging OpenAI: bring back GPT-4.5.
More from Models
- Gary Marcus says LLM math skills are like knowing only a car’s engine size — GaryMarcus · 2026-07-22
- OpenAI’s Codex + GPT-5.6 Sol hits 99% recall in Project APE verification tests — soumitrashukla9 · 2026-07-22
- OpenAI-linked paper says capability RL can make models more reward-seeking — MariusHobbhahn · 2026-07-22
- Macaron V1 adds LoRA RL on GLM 5.2 and claims SOTA benchmark gains — Xianbao_QIAN · 2026-07-22
- OpenAI rolls out voice in GPT-Live, but the UI obscures search and reasoning — Graham_dePenros · 2026-07-22
- Moonshot’s Kimi K3 sets a new open-weights ECI record at 156 — scaling01 · 2026-07-22