Artificial Analysis Accused of Tweaking Weights to Suppress Open-Source Models

Infinite-Local5435 · reddit · 2026-08-07

A user raised concerns about the impartiality of Artificial Analysis (AA) leaderboards. Previously, the open-source Qwen 3.8 max topped the platform's Agentic Index.

The user noted that AA subsequently released version "v4.1.1" of the index, adjusting the weights of benchmarks like gdpval and t3 banking to lower the open-source model's score below Anthropic's Opus. The user suspects foul play, implying the weight adjustments were made to favor proprietary models for commercial reasons.

Original post →

More from Models

Models channel →