Grok 4.7 ranks just behind Anthropic's Opus 5 on AA-Briefcase at ~50% of the cost per task

ArtificialAnlys · x · 2026-09-22

Artificial Analysis published Grok 4.7 results on AA-Briefcase, its due-diligence benchmark where models build market models and target assessment decks.

Related event: Grok 4.7 scores 46 on Intelligence Index, enters top four labs as token cost doubles(12 posts)→

Original post →

More from Models

Models channel →