Claim: "Opus 5.5" matches top human experts at picking experiments with 2.3x less compute

CatAstro_Piyush · x · 2026-10-07

Rynorhn argues science has always been bottlenecked by finite human attention — unread papers, untested hypotheses, unnoticed cross-field connections. Millions of tireless agents that read everything and run parallel experiments could turn a civilization of a few million researchers into one with effectively unlimited researchers.

Unverified claims cited: "Opus 5.5" reportedly matched the best human expert results on the TasteVal benchmark at deciding which experiments are worth running, using roughly 2.3x less experimental compute — with that ability said to double every three months.

The key shift isn't AI becoming "a better scientist" but an order-of-magnitude jump in research scale; what discovery rates look like afterward, nobody knows.

Original post →

More from AGI Musings

AGI Musings channel →