Stanford launches MMBU Challenge with Anthropic to benchmark biomedical image understanding
iScienceLuvr · x · 2026-09-12
Frontier multimodal LLMs can generate impressive answers, but reliable reasoning starts with actually understanding what's in an image — and they still struggle with biomedical visuals.
Stanford researchers, together with Anthropic, gxlai, Biohub, and the Stanford AI Lab, have launched the MMBU Challenge on the MARVL platform to benchmark and improve multimodal models' biomedical visual understanding.
More from Research
- Tao and Fields Medalists' two objections to AI in math, and why they're weak — RexDouglass · 2026-09-12
- Conjectures launches Bittensor bounties paying TAO for cracking math problems open 30-80 years, judged by machine — markjeffrey · 2026-09-12
- FADA (CoRL 2026) open-sourced: humanoid robots adapt to new conditions from 2 minutes of experience — GuanyaShi · 2026-09-12
- Intel's silicon photonics couplers hit 1-1.5 dB IL, with visible epoxy delamination flaws — jwt0625 · 2026-09-12
- CPO paper criticized for vague DLW-to-PIC coupling description: 'such as TCB' — jwt0625 · 2026-09-12
- Fly connectome trained to play a Flappy Bird–style game — TinfoilTricorn · 2026-09-12