LightOn launches LightOnOCR-3: 0.8B/4B models do OCR, captions and chart extraction in one pass
IgorCarron · x · 2026-10-09
LightOn introduced LightOnOCR-3, a family of document-intelligence models that in a single processing step recognizes text and handwriting, describes images, extracts data from charts, and preserves document structure, building on LightOnOCR-2.
On benchmarks it leads Chandra-OCR-2, Mistral OCR 4.1 and dots.mocr on olmOCR-Bench and ranks first among open-weight models on ParseBench. It ships in 0.8B and 4B versions with document processing up to twice as fast. Commenters note the chart-to-structured-data extraction is what actually matters for document pipelines.
More from Models
- Claude Haiku 5.5 jumps 257 points to 1587 on WebDev Arena leaderboard — arena · 2026-10-09
- Private visual reasoning bench eyebench saturated by Astra at 97%, author rules out data contamination — adonis_singh · 2026-10-09
- Don't crown Anthropic yet: six months ago it throttled hard, and the pendulum will swing back — ryunuck · 2026-10-09
- Haiku-5.5 reportedly hits #7 on EyeBench-v3, 76x cheaper than fable-5.1-max — adonis_singh · 2026-10-09
- ProximalHQ's training setup lifts Qwen3.8-27B to 37.2% Pass@1 on DeepSWE, up 8.4 points — aryaman2020 · 2026-10-09
- Grok moderation debate: the architecture is the culprit, not the model — xtel9 · 2026-10-09