Seven local open VLMs were tested as a second reader, and Qwen3-VL 4B won
MaziyarPanahi · x · 2026-07-22
A local evaluation of seven small open VLMs found that second-reader behavior matters more than simple text reading.
- Qwen3-VL 4B got all 7 identifiers right in 6.9s with zero false alarms and was picked for the role.
- Gemma 4 E2B reached 6/7.
- Qwen3-VL 2B was fastest but only found 4/7.
- NVIDIA LocateAnything-3B drew perfect boxes but failed to judge what was actually personal.
- LFM2.5-VL 1.6B and InternVL3 2B read the patient name correctly, then incorrectly marked it as not personal.
The takeaway: detection is easy; judgment is the hard part. The poster also links an Apache 2.0 “Reader 1” and says both models are on Hugging Face.
Related event: Qwen3-VL Beats 6 Open-Source VLMs in Medical Image Desensitization Test(3 posts)→
More from Research
- Navier-Stokes, Riemann, P vs NP: what this week's math buzzwords mean for you — koltregaskes · 2026-09-11
- Harry Collins: LLMs can't do frontier science because they can't invent new language — whoamisri · 2026-09-11
- The Waymo effect: how AI is quietly making research less collaborative — JohnHammersley · 2026-09-11
- Causal-only attention for non-generative tasks is wasteful, argues HF engineer — antoine_chaffin · 2026-09-11
- Catholic University of Chile researcher: scaling AI feedback is key to sustainable medical education — julianvarascom · 2026-09-11
- Nature paper images cellular activity across all organs, revealing body-wide circuits — arjunrajlab · 2026-09-11