DeepSeek-Vision Test: Compresses 1000 Tokens to 330

teortaxesTex · x · 2026-08-22

A test of DeepSeek-Vision's ability to process an image containing 1000 tokens (783 words) showed it compressed the input to 330 tokens while performing OCR. While efficient in token reduction, the model introduced 3 non-trivial wording changes during recognition. The user concluded it is not yet a one-shot replacement for traditional OCR.

Related event: DeepSeek-Vision Compresses 1000-Word Images to 330 Tokens with Some Errors(2 posts)→

Original post →

More from Models

Models channel →