V-CoLA: training-free vision token compression keeps 99.5% performance at half the tokens

ATH-MaaS · hf · 2026-10-09

ATH-MaaS released V-CoLA, a training-free vision token compression framework designed for linear-attention hybrid VLM architectures (e.g., Qwen3.5).

Key points:

Original post →

More from Models

Models channel →