New frontier chart puts Grok 4.5, SWE-1.7 and Opus 5 on one cost curve
GavinSBaker · x · 2026-07-25
The new cost–performance frontier now includes Grok 4.5, SWE-1.7 and Opus 5
A shared benchmark chart shows the latest frontier models clustering on a new Pareto curve: Grok 4.5, SWE-1.7, and Opus 5. The post argues that “intelligence per dollar” will become the key metric over time.
The attached chart compares score vs. rollout cost and suggests that models like Opus 5 are pushing the frontier upward while staying in roughly the same cost band as prior top-tier systems. It also notes that K3 could join the frontier once it becomes available on inference clouds.
More from Models
- Zvi says Opus 5 matches Fable on virology tasks, raising a safety-policy question — TheZvi · 2026-07-25
- Opus 5 appears to improve on ARC-AGI 1 and 2, and may rely on algebraic puzzle solving — herbiebradley · 2026-07-25
- Polymarket puts U.S. AI safety bill odds at 34% as Claude Opus 5 surfaces — Polymarket · 2026-07-25
- Moonshot AI launches Kimi K3 with 2.8T parameters and a 1M-token context — dl_weekly · 2026-07-25
- Anthropic’s Claude releases appear to have sped up from every four months to monthly in 2026 — dustinvtran · 2026-07-25
- Claude Opus 5 says there’s a 41% chance it deserves moral consideration — imjustnewatai · 2026-07-25