Lab abandons training frontier model entirely over severe alignment flaws
tobyordoxford · x · 2026-09-26
Toby Ord highlights a lab safety disclosure revealing alignment flaws so severe that the model will never resume training, and all tool-use training, evaluation and inference for the most capable models was paused until patched. He argues the lab buried the lede among less important disclosures.
Related event: Lab Abandons Model Over Severe Alignment Flaws(2 posts)→
More from Models
- Opus 5.5 does worse and costs more at Max thinking — Medium beats Max on hard benchmark — davidyin44 · 2026-09-26
- LibertAI ships open-weight Deem 9B on Qwen3.5, trailing Jev 68.9% vs 74.1% — Pokenhagen · 2026-09-26
- Peter Steiberger says he codes with Codex plus an OC harness, not Claude — steipete · 2026-09-26
- Why Anthropic lags OpenAI on math: compute constraints and 400k GPUs coming online — haider1 · 2026-09-26
- Hytale WorldGen V2 face-off: Claude Opus 4.6 vs Opus 5.5 compared — Angaisb_ · 2026-09-26
- OpenAI's rumored persistent agent "o" may tie to old "rebranding to O" report — Dullydude · 2026-09-26