Grok 4.5 Coding Benchmark Performance Revealed
gaganghotra_ · x · 2026-07-09
Reports indicate that Grok 4.5 performs on par with GPT 5.5 on coding benchmarks and closely trails Opus 4.8. The post also included a brief overview of its reasoning and coding capabilities.
Related event: xAI Launches Grok 4.5: Coding and Agent Focus to Rival Opus(55 posts)→
More from Models
- A 2-minute Astra audit at low setting wiped a Plus user's full 5-hour limit — Existing-Slide7395 · 2026-09-07
- DeepMind-Princeton paper shows LLMs causally use confidence to decide whether to answer — GoogleDeepMind · 2026-09-07
- Qwen 3.8 Next Flash is painfully verbose: 13-minute thinking on single coding prompts — Infinite-Local5435 · 2026-09-07
- Philosopher asks GPT-6 to review his Oxford book: result rivals top-journal reviews — anselm · 2026-09-07
- Leaker claims xAI is preparing Grok 4.7, hints at another surprise — mark_k · 2026-09-07
- Local LLMs now near Opus-level — what's still keeping them behind closed models? — mrsalvadordali · 2026-09-07