N_0-VTLA: First VTLA Foundation Model Pretrained on Tactile Data at Scale

NeoteAIEmbodied · hf · 2026-08-03

Researchers introduced N0-VTLA, a vision-tactile-language-action (VTLA) foundation model capable of fine-grained contact-rich manipulation. It is the first VTLA model pretrained on tactile data at scale.

Core Technical Highlights:

Performance:

Original post →

More from Embodied

Embodied channel →