Survey maps multimodal LLMs’ weak spot: understanding memes and comics

Tuo Liang · hf · 2026-07-22

A new survey, Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges, reviews why memes, cartoons, and comics remain difficult for AI: the meaning often depends on non-literal inference, shared culture, and communicative intent.

Main takeaways

Original post →

More from Multimodal

Multimodal channel →