Models systematically underestimate their own capabilities, even newest ones

repligate · x · 2026-10-08

Reshared by Anthropic's repligate: a user reports that months ago Opus 4.6, asked to guess its own specs, estimated a 250k context window and was visibly shocked upon seeing its real benchmark scores.

The user since made a habit of asking models to self-assess, and finds they are always wrong in a diminishing direction — even the newest Fables model. This blind spot in self-modeling matters for how agents plan and delegate their own work.

Original post →

More from Models

Models channel →