QWM method: morphology encoder, adaptive reward normalizer, and latent morphology conditioning
breadli428 · x · 2026-10-10
Method post of the QWM thread: Quadrupedal World Model conditions a single generative dynamics model on scale-invariant physical features and trains policies entirely inside it, via a physical morphology encoder, an adaptive reward normalizer, and latent morphology conditioning.
More from Embodied
- 20-minute explainer breaks down how Tesla trains FSD: 8 cameras, 36fps, 2B signals to 2 outputs — PTrubey · 2026-10-10
- Real Steel IRL: Humanoid Robots Noisy Boy and Raider-72 Box Each Other — zealcaiden · 2026-10-10
- Dyna launches Taku semi-humanoid robot to remove the human babysitting bottleneck — JasonMa2020 · 2026-10-10
- First-ever humanoid deathmatch was absolutely wild — cixliv · 2026-10-10
- Atomic Machines exits stealth with AI-native Matter Compiler that builds micromachines from code — DeryaTR_ · 2026-10-10
- Ex-NVIDIA GPU veterans build Physical AI chip claiming 5x performance at 80% lower cost — 创业邦 · 2026-10-10