Code OCR Test: DeepSeek Vision Compresses Tokens but Introduces Wording Errors

teortaxesTex · x · 2026-08-22

The author tested model OCR capabilities on code screenshots. DeepSeek-Vision compressed a 1000-word code screenshot into 330 tokens but introduced 3 nontrivial wording changes. The test suggests that models cannot be fully trusted for one-shot OCR replacement of high-density code screenshots unless converted to unambiguous rendering formats.

Related event: DeepSeek-Vision Compresses 1000-Word Images to 330 Tokens with Some Errors(2 posts)→

Original post →

More from Models

Models channel →