New Claude models show strong reasoning, but consume massive thinking tokens

kimmonismus · x · 2026-08-24

Users tested unreleased Claude models (suspected Opus 5.1), noting strong performance on medium reasoning tasks and 3D RL scenarios. A key observation is the heavy consumption of "thinking tokens", frequently hitting the max-tokens limit.

Original post →

More from Models

Models channel →