LightOn launches LightOnOCR-3: 0.8B/4B models do OCR, captions and chart extraction in one pass

IgorCarron · x · 2026-10-09

LightOn introduced LightOnOCR-3, a family of document-intelligence models that in a single processing step recognizes text and handwriting, describes images, extracts data from charts, and preserves document structure, building on LightOnOCR-2.

On benchmarks it leads Chandra-OCR-2, Mistral OCR 4.1 and dots.mocr on olmOCR-Bench and ranks first among open-weight models on ParseBench. It ships in 0.8B and 4B versions with document processing up to twice as fast. Commenters note the chart-to-structured-data extraction is what actually matters for document pipelines.

Related event: LightOn Open-Sources LightOnOCR-3: One Model for End-to-End Document Parsing(14 posts)→

Original post →

More from Models

Models channel →