QF3 arrives: fast flow RL trains humanoids from scratch and fine-tunes VLAs

ChongZzZhang · x · 2026-10-07

Chung Min Kim introduces QF3 (Fast Flow RL with Filtered Q-Gradients), a simple off-policy update that can train humanoids from scratch, fine-tune VLAs, and even fine-tune image models — one method covering multiple tasks. A robotics training research advance worth watching for paper details and reproductions.

Original post →

More from Embodied

Embodied channel →