Astra replication shows 1.75x reasoning steps without chain-of-thought, a 'concerning trend'
burny_tech · x · 2026-09-12
Neel Nanda replicated the Astra system card's claim of substantial computation without chain-of-thought: Astra completes 1.75x the reasoning steps of the next best models (Fable 5.1 / Gemini 3.8 Flash), with no-CoT capability jumping far more than with-CoT — 'a concerning trend'. Ryan Greenblatt adds these numbers likely underestimate the jump: ECI mishandles big jumps on saturated benchmarks, and Astra seems to unusually benefit from filler tokens.
Related event: Astra's No-CoT Reasoning Spike Raises Covert-Computation Safety Concerns(8 posts)→
More from Models
- Abacus AI Launches Smaug Flash: Open Weights at $0.10/M Input, Tuned for Personal Agents — bindureddy · 2026-09-12
- Smaug Flash launches: open-weights model tuned for personal agents at $0.10/$0.40 per M tokens — bindureddy · 2026-09-12
- "Stop Shipping Agentic Coding One-Tricks": Redditor Longs for Claude 4.5 Era Writing — warlordthe99th · 2026-09-12
- If US and Chinese labs buy training data from the same vendors, what makes models different? — biotechplays · 2026-09-12
- Grok touts its 'most intelligent model yet' with longer context and clearer reasoning — xiaosun86 · 2026-09-12
- Dev's verdict: Claude forgets tasks, Codex is too literal and adds unrequested extras — rickasaurus · 2026-09-12