Study of 18 VLMs finds answer inertia: CoT reasoning rarely revises initial predictions

delliott · x · 2026-10-06

A University of Copenhagen team (Danae Sánchez Villegas, Desmond Elliott et al.) presents at COLM an analysis of reasoning dynamics across 18 vision-language models, spanning instruction-tuned and reasoning-trained models from two families.

Key findings:

Original post →

More from Research

Research channel →