Grok 4.5 Joins Perplexity as Orchestrator, Tops WANDR Benchmark
Recently, xAI's Grok 4.5 was officially integrated into Perplexity's system, becoming an orchestrator model for Perplexity Computer's Consumer Pro/Max subscribers. This collaboration has drawn significant attention due to Grok 4.5's dominant performance in internal benchmark tests.
Benchmark Data and Performance
In Perplexity's WANDR benchmark, Grok 4.5 was compared against five other orchestrator configurations. Results showed that Grok 4.5 achieved the highest score of 0.328, making it the best-performing frontier model, with a single trial cost of approximately $4.76. Perplexity's Arav Srinivas expressed being deeply impressed by its performance, while Elon Musk emphasized that the core value of Grok Build and Grok 4.5 lies in their genuine usefulness in real-world work scenarios.
Real-World Application Feedback
In terms of actual user experience, user @BWay124 tested Grok Build powered by Grok 4.5 (High). The user noted that Grok Build has reached a completely new level, adapting to any terminal, project, or codebase environment. By simply inputting a goal and removing barriers, it can autonomously complete subsequent tasks, such as smoothly executing database modifications and queries via autonomous CLI operations for the first time. However, the user also suggested adding support for seamlessly resuming conversations across terminals or codebases to improve cross-layer workflows.
2026-07-11 ~ 2026-07-12 · 10 related posts
- Episode 1: Polymarket Bets on GPT-5.6 Release Before July 7(2026-07-03, 8 posts)
- Episode 2: GPT 5.6 Is Opus-Tier, Cheaper and Faster Than Opus 4.8(2026-07-04, 3 posts)
- Episode 3: Rumors Swirl Around OpenAI’s GPT-5.6 Launch(2026-07-05, 17 posts)
- Episode 4: Unverified Rumor Says GPT-5.6 Found New Math(2026-07-06, 2 posts)
- Episode 5: Musk Announces Grok 4.5 with 1.5T Parameters and Enhanced Coding(2026-07-07, 25 posts)
- Episode 6: Prediction Markets Strongly Price In Grok 4.4 Release(2026-07-07, 2 posts)
- Episode 7: OpenAI Announces GPT-5.6 Sol for Thursday Release Amid Early Tester Reviews(2026-07-07, 58 posts)
- Episode 8: OpenAI Launches Full-Duplex Voice Model GPT-Live(2026-07-07, 44 posts)
- Episode 9: Grok 4.5 Released with Focus on Coding and Low Cost(2026-07-08, 61 posts)
- Episode 10: New ChatGPT Voice Mode Tested: Near-Human Multi-lingual Experience(2026-07-09, 14 posts)
- Episode 11: GPT-5.6 Tested: Major Coding Leap and Direct Rival to Fable 5(2026-07-09, 30 posts)
- Episode 12: xAI Launches Grok 4.5: Coding and Agent Focus to Rival Opus(2026-07-09, 55 posts)
- Episode 13: Grok 4.5 Benchmarks Strong but Faces Data Controversy(2026-07-09, 6 posts)
- Episode 14: Rumors Swirl Over Imminent Releases of Multiple AI Models(2026-07-09, 2 posts)
- Episode 15: Grok 4.5 Receives Widespread Praise for Speed and Coding(2026-07-09, 13 posts)
- Episode 16: Grok 4.5 Praised for Impressive Speed and Performance(2026-07-09, 2 posts)
- Episode 17: Grok 4.5 Outperforms Fable in Coding Speed and Efficiency(2026-07-09, 3 posts)
- Episode 18: Grok 4.5 Released, Ranks 6th on Vals Index(2026-07-09, 2 posts)
- Episode 19: Frontier Model Comparison: GPT-5.6 Praised for Value and Creativity(2026-07-09, 3 posts)
- Episode 20: OpenAI Launches GPT-5.6 Series: Multi-Agent and Cost-Efficiency(2026-07-09, 119 posts)
- [source] Grok 4.5 Added as Orchestrator Model — perplexity_ai · 2026-07-11
- Grok 4.5 Tops Internal Benchmarks — billyuchenlin · 2026-07-11
- Grok 4.5 Scores Highest in WANDR Evaluation — AravSrinivas · 2026-07-11
- Grok 4.5 Lands on Perplexity — Kyrannio · 2026-07-11
- [source] Grok 4.5 Enters Orchestration Workflows — elonmusk · 2026-07-11
- Grok Build Tested: A Powerhouse for Terminal Tasks — BWay124 · 2026-07-11
- [source] Grok Build Tested: Handles Database Operations — BWay124 · 2026-07-11
- Grok 4.5 Enabled on Perplexity Enterprise — testingcatalog · 2026-07-11
- Grok 4.5 Tops as Best Orchestrator — rohanpaul_ai · 2026-07-12
1 near-duplicate retellings: rohanpaul_ai