MonkeyOCRv2: Document AI Vision Model

VLRLab-OCR · hf · 2026-07-15

This paper introduces MonkeyOCRv2, a vision-text foundation model designed for Document AI.

Key highlights include:

Related event: MonkeyOCRv2: A New Document AI Foundation Model(2 posts)→

Original post →

More from Multimodal

Multimodal channel →