Leaked Grok 4.6 Leads in Agent Tasks, Beats GPT-5.6 in Coding
rohanpaul_ai · x · 2026-08-13
Rumors suggest xAI released Grok 4.6. Its reported standout feature is professional agent work, leading GPT Sol Max and Fable 5 Max on GDPVal-AA v2 and AA-Briefcase. It also matches GPT-5.6 at 61 on Artificial Analysis while costing $2/$6 per million tokens. For coding, it purportedly scores 69.9% on CursorBench v3.2, beating Sol's 67.2%. Note: Model names and data are likely unverified or fictional.
Related event: xAI Releases Grok 4.6: Top-Tier Performance at Unmatched Cost Efficiency(77 posts)→
More from Fun
- fofrAI Shares AI-Generated Thriller Clip: Bystander's Reaction Steals the Show — fofrAI · 2026-08-13
- AI Creates the Most Disrespectful Boss Dodge in Gaming History — heypearlai · 2026-08-13
- Creepy or Cool? AI Ads Could Shift Features to Match Your Gaze — ZeroStateReflex · 2026-08-13
- AI Benchmark Parody: Fake Claude Opus 5 Sweeps InferenceBench Top 5 — maksym_andr · 2026-08-13
- Joke: Netizen Suggests Anthropic Release 'Marlboro White Monster' to Fix PR — tekbog · 2026-08-13
- User Reflects: We Really Live in a Cyberpunk Reality — RileyRalmuto · 2026-08-13