0.8B Qwen 3.5 full finetune transcribes 2,000+ medieval manuscripts across 10,000 pages

wjb_mattingly · x · 2026-09-04

wjbmattingly shares a full finetune of 0.8B Qwen 3.5 trained on 2,000+ medieval manuscripts (10,000 pages total) from the Comma dataset (0.09 CER), to be presented with models in Vienna. He acknowledges CER is no longer a good default eval and plans to compare other approaches.

Related event: 0.8B Qwen Fine-Tuned on Medieval Manuscripts Sparks CER Debate(4 posts)→

Original post →

More from Research

Research channel →