VLAct: Representation-Centric Continued Pre-training for Vision-Language-Action Models

StarVLA · hf · 2026-08-31

VLAct improves VLA performance via representation-centric continued pre-training on diverse robot data with preserved vision-language priors and shared action semantics, achieving strong results with limited compute.

Original post →

More from Embodied

Embodied channel →