AI IQ Launches: Frontier Models Like GPT-5.5 and Opus 4.7 Scored on Human IQ Scale
beffjezos · x · 2026-09-06
Ryan E. Shea launched AI IQ, a benchmark scoring frontier models — GPT-5.5, Claude Opus 4.7, Gemini 3.1, Grok 4.3, Kimi K2.6, Qwen3.6, DeepSeek V4, Muse Spark and more — on the human IQ bell curve, tracking frontier IQ over time, comparing IQ vs EQ, and quantifying what intelligence costs in practice. beffjezos asked for a GPT-6 Astra test.
More from Models
- Unverified: GPT-6 Astra reportedly scores 91.8% on SpatialBench vs 80% human baseline — VraserX · 2026-09-06
- DeepSeek v4 models reportedly 30% off via third-party channel, unverified — matlabulous · 2026-09-06
- One prompt reportedly got GPT-6 Astra to build a full game and record a video — yungcontent · 2026-09-06
- One year of model progress: from broken sprite animations to fully rigged playable games — Dimillian · 2026-09-06
- AtCoder Founder's Private Bench: Astra Leads Fable5.1 by Roughly Two Generations — i_dg23 · 2026-09-06
- Ollama CEO: open models will carry 80-90% of enterprise tokens at just 10-20% of cost — victor_explore · 2026-09-06