Stanford Benchmarks 32 Foundation Models on Pathology: Vision Models Outperform Vision-Language
iScienceLuvr · x · 2026-08-09
Stanford researchers evaluated 32 foundation models across a diverse suite of 41 pathology tasks. Key findings indicate that pathology-specific vision models (Path-VM) outperform pathology-specific vision-language models (Path-VLM). Additionally, simply scaling model size and training data does not uniformly improve pathology performance, while model ensembling effectively boosts task accuracy.
The top-performing individual models on the benchmark include Virchow2 and Prov-GigaPath (Microsoft), UNI (Harvard), and H-Optimus-0 (Bioptimus).
More from Research
- Retriever: A Framework for Asynchronous, Closed-Loop Robot Agents — ZeYanjie · 2026-08-24
- Converting GMMs ↔ PEFs for fast KLD approximation — FrnkNlsn · 2026-08-24
- Netflix details its production LLM judge: hundreds of thousands of recommendations scored weekly — omarsar0 · 2026-08-24
- Nature Comment: Provenance, not interpretability, grounds trust in autonomous science — gabepgomes · 2026-08-24
- New Architecture RHEA: Train 1B Model on 8GB VRAM — zemondza · 2026-08-24
- Trained two 16M-param models to do generative CAD with real physics — debreuil · 2026-08-24