COLM 2026 paper maps how much VLM representations leak across layers

sineadwilliamso · x · 2026-10-07

Presenting 'What do your logits know' at COLM 2026, the authors systematically compared information retained at different representational levels of Vision-Language Models — from the rich residual stream through two natural bottlenecks: tuned lens projections and final top-k logits. The core question: when you ask a yes/no question about an image, how much other information leaks through the output? Far more than you might expect.

Related event: Study: VLM top-k logits leak far more task-irrelevant image information than expected(8 posts)→

Original post →

More from Research

Research channel →