SECRET: training-free relay steering cuts cross-modal hallucinations in AVLLMs by up to 18 points

Yu Zhang · hf · 2026-10-02

This paper dissects source-confused grounding hallucinations in audio-visual LLMs via path-intervention analysis, revealing a question-relay mechanism where interfering modality cues contaminate question states. It proposes SECRET, a training-free steering method that outperforms prior training-free baselines on CMM and AVHBench across three AVLLMs, with gains up to +18.0 percentage points.

Original post →

More from Research

Research channel →