NUS's RL² Framework Boosts VLA Out-of-Domain Success Rates by 17%

NationalUniversityofSingapore · hf · 2026-08-03

National University of Singapore introduced RL², an adaptive inference-time steering framework that trains a lightweight offline RL policy on the latents of Vision-Language-Action (VLA) models.

Original post →

More from Embodied

Embodied channel →