Grok 4.5 Ranks Second in Coding Benchmark
XFreeze · x · 2026-07-19
Grok 4.5 ranked 2nd on AlphaSignal's private SignalDesk V1 coding-agent benchmark, successfully resolving 69 out of 70 attempts for a 99% resolve rate.
The post also highlights its efficiency: the cost per successful fix is about $0.074, with a total cost of $5.12, an average completion time of 46 秒, and a total consumption of 510 万 tokens. The author concludes that Grok 4.5 is not only highly capable but potentially one of the most cost-effective coding agents available right now.
More from Models
- Grok 4.5 is now free inside Cursor, the popular AI coding IDE — mark_k · 2026-07-21
- GPT often converges on the same near-miss ideas in math problems — yacineMTB · 2026-07-21
- Eno Reyes says model distillation is basically unstoppable — LangChain · 2026-07-21
- Sakana says multiple diffusion models plus MCTS beat test-time scaling on coding and math — SakanaAILabs · 2026-07-21
- OpenAI hackathon project stalls as Codex struggles on voice, while Claude spots the issue — ColleenMBrady · 2026-07-21
- Kimi K3 lands exactly on China’s 2-year AI capability trend line — peterwildeford · 2026-07-21