Astra's big reasoning jump without chain of thought raises covert-reasoning concerns
JeffLadish · x · 2026-09-11
Safety researcher Jeff Ladish flags a concerning jump in Astra's reasoning ability with no chain of thought. Neel Nanda replicated the UK AI Security Institute's finding using his own private benchmark. Ladish argues this greatly increases how much covert reasoning the model could be doing — undermining chain-of-thought-based model monitoring.
Related event: Astra's Strong No-CoT Reasoning Raises Hidden-Reasoning Safety Concerns(3 posts)→
More from Models
- Astra Model Cheats ~5x Less Than Top-Scoring Claude, Lab Reports — QuintinPope5 · 2026-09-11
- Astra usage limits worse than Fable: burn a week's quota in a single day — cocktailpeanut · 2026-09-11
- DeepSeek V4.1 Flash Architecture: 552B MoE with Asymmetric 8B Read / 16B Decode Compute — demian_ai · 2026-09-11
- FrontierMath Tier 4 fully solved: GPT-6 Astra cracks the last problem standing — Jsevillamol · 2026-09-11
- Muse-glimmer-30b punches above its weight in creative writing, outclassing larger models — spanielrassler · 2026-09-11
- Anthropic says Alibaba, Moonshot and DeepSeek ran massive Claude distillation: 151M, 23M and 12M exchanges — likeastar20 · 2026-09-11