qinglong-captions v4.7.0: Major Upgrades to Multimodal OCR and Underlying Models

bdsqlsz · x · 2026-08-02

The image captioning tool qinglong-captions has released v4.7.0, bringing several updates to multimodal capabilities and model support:

Related event: qinglong-captions v4.7.0 Adds Multimodal OCR and Score Recognition(2 posts)→

Original post →

More from Multimodal

Multimodal channel →