ECCV 2026 paper says some multimodal LLMs do worse with one-shot examples

HildeKuehne · x · 2026-07-29

Some multimodal LLMs get worse with one-shot examples, and the paper says class names are part of the reason

The post argues that on standard few-shot benchmarks, several MLLMs do better 0-shot than 1-shot. When semantic class names are removed, in-context prompting can nearly collapse.

The linked ECCV 2026 paper proposes DeCoDe and claims it explains why examples can hurt instead of help in some multimodal settings.

Original post →

More from Research

Research channel →