Dirty Tokens Cause Model Anomalies Linked to Training Pathologies

teortaxesTex · x · 2026-08-21

Certain substrings that most models process fine can trigger anomalous behaviors in others. This is highly specific per model/tokenizer and linked to pathologies in training tokens matching those substrings.

Related event: Researchers Find Abnormal Tokens Causing Weird DeepSeek-V3 Outputs(2 posts)→

Original post →

More from Models

Models channel →