Tsinghua's KBMR embeds images by semantic entity, boosting knowledge-based visual QA

Tsinghua · hf · 2026-09-03

Tsinghua researchers introduced KBMR, a retrieval method for knowledge-based visual question answering.

Original post →

More from Multimodal

Multimodal channel →