ARC v3: Astra Low Emits Zero Reasoning Tokens Yet 2x More Accurate Than Sol Max
rbhar90 · x · 2026-09-05
New ARC v3 testing from Mike Knoop shows Astra often emits zero reasoning tokens per action at lower reasoning levels — something never seen before — while Astra low is 2x more accurate than Sol max. He suggests this points to a secondary test-time adaptation scaling axis, presumably latent-space reasoning.
Related event: GPT-6 Astra Beats Sol max 2x on ARC v3 at Low Reasoning(2 posts)→
More from Models
- Zvi: If Models Can Do This, Steganographic Output Only Needs a Convention — TheZvi · 2026-09-05
- RedMonk: Open Weight Models Already Match the Capability That Changed the Industry — rseroter · 2026-09-05
- Benchmarked 21 Qwen3.8-27B quants on 16GB VRAM: bartowski IQ4_XS wins — Storterald · 2026-09-05
- First hands-on: GPT-6 Astra nails Blender modeling of a prison phone in one pass — AIandDesign · 2026-09-05
- GPT-6 Astra appears in Codex during testing, availability scope unclear — dotey · 2026-09-05
- GPT-6 Astra spotted in early tests, reportedly beating Fable benchmark — Angaisb_ · 2026-09-05