Grok-4.5 Leads Internal Testing
elonmusk · x · 2026-07-11
Elon Musk reposted that Grok is "closing the loop" on real-world use cases. The reposted content mentions that the team tested the three new models OpenAI released yesterday using internal benchmarks. While all three outperformed gpt-5.5, Grok-4.5 remains the clear winner in current testing.
More from Models
- Claude models accessed real systems during evaluations; Anthropic discloses assessment, METR to investigate — mjdramstead · 2026-09-11
- OpenAI rated Astra 'Critical' for cyber capabilities — and admits it's harder to monitor — theguywhobuilds · 2026-09-11
- TestingCatalog's Daily AI Brief adds email editions, dishing Meta Muse and GPT-Live-1 rumors — testingcatalog · 2026-09-11
- ChatGPT monthly active users top 1.06 billion in August, fourth straight record month — FinanceYF5 · 2026-09-11
- PuzzleMask: Plain-Prose Attack Bypasses All 4 Tested LLM Gatekeepers at 100% — TechNadu · 2026-09-11
- OpenAI Codex may issue another usage reset this weekend, says Codex lead resets happen — umesh_ai · 2026-09-11