AutoPrune: LLMs Automatically Design Visual Token Pruning for Multimodal Models

Zhen Liu · hf · 2026-08-14

AutoPrune introduces an "AI4AI" framework that leverages large language models (LLMs) to automatically design visual-token pruning policies for multimodal models.

The method utilizes a domain-specific language (DSL) and a residual search formulation to generate and optimize these policies. Its core objective is to significantly boost the inference efficiency of multimodal models while minimizing performance degradation during token pruning.

Original post →

More from Research

Research channel →