Grok-4.5 Leads Internal Testing
elonmusk · x · 2026-07-11
Elon Musk reposted that Grok is "closing the loop" on real-world use cases. The reposted content mentions that the team tested the three new models OpenAI released yesterday using internal benchmarks. While all three outperformed gpt-5.5, Grok-4.5 remains the clear winner in current testing.
More from Models
- Google says information agents are coming to AI Pro and Ultra this summer — gaganghotra_ · 2026-07-22
- Poolside’s Laguna S 2.1 gets a two-week free run on Nous Portal — NousResearch · 2026-07-22
- Qwen3.8 Max Preview looks substantially better in a side-by-side test with Kimi K3 — curiousily_ · 2026-07-22
- Moonshot’s Kimi K3 reaches #5 on MathArena as the top open model — xeophon · 2026-07-22
- Google launches Gemini 3.5 Flash Cyber for CodeMender, with limited access for governments — GoogleAI · 2026-07-22
- Gemini 3.5 Flash-Lite beats 3.1 Flash-Lite on long-context retrieval in MRCRv2 — Dillonu · 2026-07-22