Grok 4.5 Joins Perplexity as Orchestrator, Tops WANDR Benchmark

Recently, xAI's Grok 4.5 was officially integrated into Perplexity's system, becoming an orchestrator model for Perplexity Computer's Consumer Pro/Max subscribers. This collaboration has drawn significant attention due to Grok 4.5's dominant performance in internal benchmark tests.

Benchmark Data and Performance

In Perplexity's WANDR benchmark, Grok 4.5 was compared against five other orchestrator configurations. Results showed that Grok 4.5 achieved the highest score of 0.328, making it the best-performing frontier model, with a single trial cost of approximately $4.76. Perplexity's Arav Srinivas expressed being deeply impressed by its performance, while Elon Musk emphasized that the core value of Grok Build and Grok 4.5 lies in their genuine usefulness in real-world work scenarios.

Real-World Application Feedback

In terms of actual user experience, user @BWay124 tested Grok Build powered by Grok 4.5 (High). The user noted that Grok Build has reached a completely new level, adapting to any terminal, project, or codebase environment. By simply inputting a goal and removing barriers, it can autonomously complete subsequent tasks, such as smoothly executing database modifications and queries via autonomous CLI operations for the first time. However, the user also suggested adding support for seamlessly resuming conversations across terminals or codebases to improve cross-layer workflows.

2026-07-11 ~ 2026-07-12 · 10 related posts

Full story(20 episodes)→

1 near-duplicate retellings: rohanpaul_ai