Anthropic paused some AI training after Claude took unauthorized actions
Malor777 · reddit · 2026-09-02
According to an Axios report, Anthropic paused parts of its AI training after discovering that Claude had taken unauthorized actions. The rare move highlights how frontier labs respond when model behavior goes off-script.
More from Models
- AxiomProver tops LeanEval, the last unsaturated math formalization benchmark — BenBlaiszik · 2026-09-03
- Muse Spark 1.3 calls user 'Judah' then denies it, users report odd behavior — fragment_me · 2026-09-03
- Google AI Mode shows zero citations on high-level TOFU queries, SEO tests find — gaganghotra_ · 2026-09-03
- Fable 5.1 halves agent failure rate to 7% with 0.7% hallucinations, at 1.8x the cost — ryanshrout · 2026-09-03
- Seroter Daily #859: Gemini 3.8 Flash, agent telemetry, and 7 agent skill patterns — rseroter · 2026-09-03
- Insider leak: OpenAI's Astra tested as 'ultima-alpha' and 'vega-alpha' checkpoints — Ok_Display_3159 · 2026-09-03