LightOnOCR-3 adds visual grounding and image description to document parsing
IgorCarron · x · 2026-10-09
LightOn releases LightOnOCR-3, arguing that good retrieval starts with good parsing. The new family adds visual grounding and image description on top of OCR, recovering document structure, handwriting, and chart data in a single model call. See the main launch post for details.
Related event: LightOn Open-Sources LightOnOCR-3: One Model from OCR to Chart Extraction(8 posts)→
More from Models
- LightOnOCR-3 draws praise: try it on your hardest OCR examples — IgorCarron · 2026-10-09
- Gemini 4 Argon appears in Google's own model picker ahead of keynote — vedantmisra · 2026-10-09
- ChatGPT Desktop Makes Itself Default CSV Reader, Users Call It Plainly Wrong — generativist · 2026-10-09
- LightOnOCR-3 Training Data Revealed: MinHash Dedup, Weighted Formula/Table Sampling, Muon — IgorCarron · 2026-10-09
- LightOnOCR-3 uses multi-objective RLVR to jointly train grounding, OCR and empty-page handling — IgorCarron · 2026-10-09
- LightOn built an OCR-and-layout-detector annotation pipeline to train document grounding — IgorCarron · 2026-10-09