Moonshot Criticized for Ignoring ARC-AGI Benchmark, Focusing on Coding and Writing

bookwormengr · x · 2026-08-01

A user criticized Moonshot for neglecting challenging reasoning benchmarks like ARC-AGI, arguing the company only focuses on coding, GDPEval, and creative writing. Citing OpenAI as an example, the post noted that while early GPT-5-Pro struggled with ARC-AGI, later versions improved significantly, urging frontier labs to prioritize such revolutionary evaluations. The original post also complained about Kimi's current pricing lacking competitiveness.

Original post →

More from Models

Models channel →