Elon Musk Announces Grok 4.5: Beats GPT-5.6 in Benchmarks, Launches CLI Coding Agent
elonmusk · x · 2026-08-01
Elon Musk officially announced the release of Grok 4.5. According to benchmark data, the new model outperforms GPT-5.6 Terra across almost all shared evaluations, including ACB, GPQA, SWE-P, and Atlas.
Alongside the model, xAI introduced Grok Build, a terminal-based coding agent powered by Grok 4.5. Key features include:
- Skills & Plugins: Supports AGENTS.md, plugins, hooks, and MCP servers out of the box. Users can capture any session as a new skill via /skillify.
- Plan Mode: For complex tasks, the agent generates a structured plan that requires user approval before executing code changes, providing a clean diff for review.
- Marketplace: Allows installing team tool bundles from a central marketplace or self-hosted Git repos, integrating seamlessly with services like Linear, Sentry, and Postgres.
More from coding & agent
- Month of AI Bugs Returns: Over 20 AI System Vulnerabilities to Be Disclosed — wunderwuzzi23 · 2026-08-01
- Meta Engineer Shares MLSys Keynote: Using AI to Liberate Systems Researchers — salykova_ · 2026-08-01
- React Aria launches TokenField component for building AI prompt inputs — pacocoursey · 2026-08-01
- Supabase Launches Evals to Benchmark AI Coding Agents on Real Tasks — tristanbob · 2026-08-01
- Building Software While Sleeping: 3 Lessons from Managing Autonomous AI Employees — leebase65 · 2026-08-01
- Kanbots: Run 11 AI Coding Agents in Parallel on One Kanban Board — tom_doerr · 2026-08-01