SeeQ ablations: TD learning, VLM pretraining and language subtask decoding all matter

aviral_kumar2 · x · 2026-09-29

Aviral Kumar shares SeeQ ablation results highlighting three key ingredients: TD learning done the best-of-N way; both pretrained VLMs and pretraining on robot data; and actually decoding the active subtask in language. Further diagnostics of the learned value function are in the paper.

Related event: CMU's SeeQ: Generalist Robot Value Functions Nearly Double Success Rates(3 posts)→

Original post →

More from Embodied

Embodied channel →