Frontier Models See Intensive Updates in 8 Days
ArtificialAnlys · x · 2026-07-18
Artificial Analysis noted a flurry of frontier model updates within 8 days: Grok 4.5, GPT-5.6, Muse Spark 1.1, and Kimi K3 dropping successively.
Key takeaways:
- 6 labs now have models scoring over 50 on the Artificial Analysis Intelligence Index, a significant jump from just 2 in early June.
- Kimi K3 entered the overall leaderboard at No. 3 with a score of 57. It excels in agentic and knowledge work capabilities, rivaling or beating some more expensive models across multiple benchmarks.
- Inference costs for near-frontier models dropped 2–3x in a single week. The article highlights that per-task costs for GPT-5.6, Grok 4.5, Muse Spark 1.1, and Kimi K3 are significantly lower than peer models from just days ago.
- Claude Fable 5 remains No. 1, but its lead is shrinking, and the price of "underlying intelligence" continues to drop.
Related event: Artificial Analysis Says Four Frontier Models Launched in Eight Days(5 posts)→
More from Infra
- Nvidia Is Now Core to Every Major Robotaxi Stack at Commercial Scale — pdamodaran · 2026-09-11
- 12 KV Cache Reduction Techniques Every AI Engineer Should Understand, Explained — blaizedsouza · 2026-09-11
- The shadow GPU capacity market is formalizing, with Meta selling excess compute to outside buyers — DavidLinthicum · 2026-09-11
- Engram's random reads don't suit SSDs; CPU-memory over NVLink could serve all 72 GPUs — bookwormengr · 2026-09-11
- 80% of the DIY LLM inference hype posters have already quit — it's brutally hard systems work — abhijithneil · 2026-09-11
- Hugging Face's Ultra Scale Playbook: a free book on training LLMs on GPU clusters — mdancho84 · 2026-09-11