Claude Opus 5 tops LisanBench while using far fewer tokens in medium mode

scaling01 · x · 2026-07-29

A post claims Claude Opus 5 is performing strongly on LisanBench. It says Opus 5 high is now the overall #1, but uses far more tokens than Opus 4.8 at high effort.

It also claims Opus 5 medium is nearly as good as Opus 4.8 high while using about half the tokens, and that Opus 5 without thinking still scores almost as high as GPT-5-medium, which reportedly used over 30k reasoning tokens on average versus 510 for Opus 5.

Original post →

More from Models

Models channel →