Apple Open-Sources LensVLM-9B, a Qwen3.5-9B Finetune That Shrinks Long Docs into Page Images
victormustar · x · 2026-09-24
Apple quietly released LensVLM-9B on Hugging Face, a finetune of Qwen3.5-9B that turns long documents into small page images to save tokens, then retrieves the full text of only the pages relevant to a user's question.
The page-image retrieval approach offers a novel token-efficient path for long-document understanding, and the model is open for download.
Related event: Apple Open-Sources LensVLM-9B to Save Tokens via Compressed Document Images(2 posts)→
More from Models
- OpenAI launches MentalHealthBench with 80+ clinicians; replies turn into memes — Yuchenj_UW · 2026-09-24
- A 10-year trend holds: small fine-tuned models on selective data still beat bigger general models — xeophon · 2026-09-24
- GPT-6 Luna shows vision regression vs GPT-5.6: extraction drops 81.79% to 66.67% — ducha_aiki · 2026-09-24
- Luna 6 private coding evals don't look great — Maasu · 2026-09-24
- Claude Opus 5.5 tops Artificial Analysis at 58; four new models add 11 Pareto frontier points — ArtificialAnlys · 2026-09-24
- METR says it used an undisclosed 'additional source' to understand Anthropic's AI R&D, buried in the Opus 5.5 system card — coherence · 2026-09-24