Mystery model "Astra" reportedly beats 5.6 Sol Pro on FrontierMath T4
ctjlewis · x · 2026-09-04
X user FateOfMuffins flags to AI safety researcher Ryan Greenblatt that a mysterious model labeled "Astra (NONE)" scores higher than "5.6 Sol Pro (Max)" on the FrontierMath T4 benchmark. The reposter reacts with "they fucking cooked." The model's identity is unknown and the claim is unverified.
More from Models
- GPT-6 Astra Tops ValsAI Code Migration Benchmark at 68% Accuracy, 2-4x Faster — sandersted · 2026-09-04
- OpenAI Launches GPT-6 Astra: An Agent That Can Do Anything on Your Computer — astralmatrix · 2026-09-04
- Andrew Ng: We need harder evals — have frontier models chat with me — andrewgwils · 2026-09-04
- Epoch AI launches FrontierMath Erdős benchmark of 68 unsolved problems; Astra tops it at 2/68 — keviv9 · 2026-09-04
- OpenAI Launches GPT-6 Astra: 99.9% on ARC-AGI-3, but Independent Evals Call It Uneven — Latent Space · 2026-09-04
- Liquid AI launches Nanos: task-specific 350M-2.6B models that run on-device — JosephJacks_ · 2026-09-04