Comparing tokens across different models is meaningless, metrics need refinement in the reasoning era

adamdangelo · x · 2026-08-27

Adam D'Angelo argues that aggregating tokens from multiple models is meaningless because tokens cannot be compared across models. This has always been true but becomes increasingly critical as models diversify and reasoning tokens constitute a larger portion of usage.

Original post →

More from Models

Models channel →