Astra Solves Mensa Puzzle on Second Try While Sol Fails Twice in Reddit Test
kaljakin · reddit · 2026-09-07
Reddit user kaljakin compared two models on Mensa-style puzzles: Astra Extra High failed its first attempt but got it right on a second try in under a minute, while Sol High (5.6) failed both attempts with two minutes of thinking time each. The author plans to track Astra's official IQ score on TrackingAI.org and notes his benchmark is still not saturated.
More from Models
- Ask LLMs to pick a random number 1-30: most all answer 17 — ohnag_eryeah · 2026-09-07
- Polymarket puts 82% odds on Anthropic releasing next Claude Opus this month — Polymarket · 2026-09-07
- Early test: GPT-6 Astra beats GPT-5.6 Sol while using fewer tokens — haider1 · 2026-09-07
- Viral screenshot shows Astra 6 at xhigh responding with 'agi' — viksit · 2026-09-07
- OpenAI Employee Predicts Major Blender x Astra Gains by Mid-2027 — Distinct-Question-16 · 2026-09-07
- GPT-6 Astra builds exploded view of 71,492-atom GPCR simulation in membrane — DeryaTR_ · 2026-09-07