Dawn Robotics Open-Sources Puffin-World, a World Model Grounded in World State
jiqizhixin · x · 2026-09-20
Dawn Robotics and NTU S-Lab present Puffin-World, a unified multimodal world model natively grounded in world state.
Key idea: for robot training, visually realistic video alone is insufficient — you need environment information supporting perception, localization, planning and action: current camera pose, stable gravity direction, continuous scene geometry, and the next view after movement. Puffin-World goes beyond relative viewpoint changes to include absolute camera orientation relative to gravity and the real world.
The team open-sources Puffin-16M: 15M vision-language-camera triples, 1M challenging rotation trajectories, and absolute camera pose annotations for 44.5M frames.
More from Embodied
- "Everyone: slow down AI. Me: what if we gave it fingers" robot meme — GregCook2011 · 2026-09-20
- Snap Specs AR glasses integrate with Salesforce Agentforce, unveiled at Dreamforce — jevon · 2026-09-20
- French Cancer Institute Deploys Mirokaï Robot to Comfort Kids in Radiotherapy — CyberRobooo · 2026-09-20
- What are we actually seeing with GPT-6 Astra? A skeptic's look at the robotics results — GeorgiaChal · 2026-09-19
- Apple M6 bumps cores to 12 with two super cores; CPUs keep improving fast — lemire · 2026-09-19
- NVIDIA Opens 2027 Applied Research Internships on Humanoid Robots (Isaac Loco-Manipulation) — YuXiang_IRVL · 2026-09-19