Intelligence Index v4.3: Claude Fable 5.1, Muse Spark 1.3 and GPT-6 Astra Reset the Cost-Efficiency Frontier
ArtificialAnlys · x · 2026-09-10
- Intelligence Index v4.3 released: replaces τ³-Banking with AutomationBench-AA, upgrades Terminal-Bench to v4.0.
- Last week the intelligence-vs-cost Pareto frontier moved out substantially — Claude Fable 5.1, Muse Spark 1.3, and GPT-6 Astra each set new efficiency points; Meta's Muse Spark 1.3 is judged to have "reached the frontier."
- New tools launched: Search Index comparing search API providers, Optima custom benchmark builder, and a model recommender.
- Recent evaluations include MiniCPM5-2B, K2 Horizon 375B A23B, and Google's Gemini 3.8 Flash — its fourth Flash model in under four months.
More from Models
- Codex CLI 0.154.0 ships GPT-6-Astra, experimental worktree support — github-actions[bot] · 2026-09-10
- OpenAI's new model, in training since Aug 28, reportedly beat GPT-6-Astra in just one week — alexcovo_eth · 2026-09-10
- Qwen3.8-2.4T-A95B open weights land on AWS: single 8×B300 node with vLLM — AWS ML Blog · 2026-09-10
- Users say Astra's $200 sub is no longer enough: multi-project work burns through quota in days — CtrlAltDwayne · 2026-09-10
- OpenAI launches GPT-6 Astra to power ChatGPT Work with desktop app control — OpenAI · 2026-09-10
- Claude Fable 5.1 cuts agreement openers 58% and em dashes 32%, Arena analysis finds — rohanpaul_ai · 2026-09-10