Unlimited-OCR reads 100-page PDFs in one shot

Shruti_0810 · x · 2026-07-20

China has open-sourced **Unlimited-OCR**, a 3B model for reading full documents in one pass. - Uses a **32K context window** to process an entire document at once instead of page by page - Preserves **cross-page context**, which helps with tables, references, and long-form documents - Reports **93%** on standard OCR benchmarks, about **+6** over the baseline - Claims **<0.11 error rate** beyond 40 pages - Runs **fully locally** and supports **Transformers, vLLM, Ollama, Docker, llama.cpp**, and more The post argues this approach makes enterprise OCR more reliable while eliminating per-page cloud OCR costs.

Related event: Baidu Open-Sources Unlimited-OCR for Long Documents(4 posts)→

Original post →

More from Multimodal

Multimodal channel →