Why models miscount letters in "strawberry": they see tokens, not letters

Euphoric_North_745 · reddit · 2026-09-28

A Reddit post explains the root of the classic "how many R's in strawberry" failure: the ChatGPT app shows your typed text, but the API actually receives token ID sequences (e.g., [5299, 1991, 460, 382, 101830, 30]).

The model reasons in token numbers, not human letters — there are no letters in its "brain" — and its numeric output is converted back to text by the API. A translator sits between you and the model, which is why counting letters trips it up. The post links to OpenAI's tokenizer tool for verification.

Original post →

More from Models

Models channel →