VASC: training-free sparse attention speeds up 3D reconstruction inference by up to 2.29x

zhenjun_zhao · x · 2026-10-02

A new arXiv paper introduces VASC (Value-Aware Sparse Attention with Cross-Layer Memory), a training-free method addressing the quadratic cost of global attention in feed-forward 3D vision models like VGGT:

On 7Scenes and NeuralRGB-D with VGGT and π³, it beats FasterVGGT on pose estimation and reconstruction quality, with up to 2.29× faster inference than dense VGGT. Code is released.

Original post →

More from Research

Research channel →