Opus 5.5 shows 2.3x the research taste of top human experts at ~1/30th the cost

SaxenaNayan · x · 2026-10-07

Periodic Labs reports Opus 5.5 has 2.3x the "experimental research taste" of its best human experts: on a typical task it matches the expert's best score in 17 GPU hours instead of 40. Cost gap is starker — an Opus 5.5 attempt costs $280 (API + GPU) vs $9,123 for an expert attempt, roughly 30x cheaper. Frontier models' experimental research taste has doubled every 3 months since December 2025.

Related event: AI's Ability to Predict Experiments Is Surging, Tests Show(2 posts)→

Original post →

More from Models

Models channel →