Baidu open-sources Unlimited-OCR, a 3B model that reads 40-page documents in one pass
HankYeomans · x · 2026-07-24
A thread says Baidu has open-sourced Unlimited-OCR, an OCR model designed to read entire documents in one pass.
- The model has 3B parameters, with only 500M active at inference time.
- It can run 100% locally and uses a 32K context window.
- The goal is to preserve text, formulas, tables, and reading order across pages, instead of chopping documents page by page.
- The thread claims 93% accuracy on a standard benchmark, about 6 points above the baseline, and an error rate below 0.11 even beyond 40 pages.
Related event: Baidu Open-Sources Unlimited-OCR for Long Documents(3 posts)→
More from Models
- Grok and Claude get personified as Elon and Dario in a new model-mood meme — kevinnbass · 2026-07-24
- Kimi K3 beats GLM 5.2 on 100 deep-research tasks, but costs 5x more — AravSrinivas · 2026-07-24
- Kimi K3 and Claude Fable5 get called the best large-model frontend aesthetes — vista8 · 2026-07-24
- Celeris-1 launches with diffusion inference, 157 ms latency and 76% MMLU-Pro — timshi_ai · 2026-07-24
- Dev rebuilds interactive 3D globe in 1.5 hours using Kimi K3 — DuRuofei · 2026-07-24
- A Kimi demo claims it rebuilt a Google Maps 3D experience in 1.5 hours — shakoistsLog · 2026-07-24