HiDream launches embodied world model HiDream-O1-Embodied, tops RoboColiseum robustness board at 0.692
机器之心 · wechat · 2026-09-07
HiDream.ai released HiDream-O1-Embodied, an embodied world model aimed at turning video generation into action control with stronger physical perception and dynamic prediction. On its first RoboColiseum evaluation, the model topped the Robustness sub-board with an average score of 0.692. The platform runs 78 high-fidelity simulation tasks, testing stability under changed lighting, materials, camera positions and rewritten instructions.
Key technical claims include intent-level language understanding beyond keyword matching, multi-view visual perception that keeps running when some views fail, and fault-tolerant training with deliberately degraded conditions. Data-wise, HiDream uses a "real base + generative augmentation" paradigm, expanding mocap data 100x while preserving physical constraints.
CTO Yao Ting said full-modality representation, causal reasoning and physical world modeling together form a complete world model foundation. A month earlier HiDream topped the WBench Navi board at 80.9 with its interactive world model HiDream-O1-World.
More from Embodied
- KinetixAI's humanoid KAI looks human: 1.73m, 70kg, 115 DoF, tactile skin — CyberRobooo · 2026-09-07
- Xiaomi's CyberOne robot hits IFA as European consumer tech brands go missing — alexmacgregor__ · 2026-09-07
- Dev predicts 10k-100k tok/s inference will make frontier models real-time robot policies — eigenron · 2026-09-07
- Astra robot-control demos pile up; speculation OpenAI is building its own Gemini Robotics rival — CyberRobooo · 2026-09-07
- GPT-6 "Astra" shows zero-shot robot arm sorting; OpenAI robotics rumors swirl — CyberRobooo · 2026-09-07
- HiSfM: scaffold-anchored hierarchical SfM tames repeated-structure ambiguity and cuts runtime — zhenjun_zhao · 2026-09-07