Hyper3D launches WorldGen: one image to a full editable 3D scene
智东西 · wechat · 2026-09-01
Shanghai-based Hyper3D (Yingmou) released WorldGen, a world-generation model that moves 3D generation from single assets to complete scenes. From one scene photo, it auto-detects major objects (2-3s per-object preview, full scene in 2-3 minutes) and produces multiple independent, editable, interactive assets. Its core CAST architecture won a SIGGRAPH 2025 Best Paper award, recovering occluded structure from a single image and inferring contact/support relations between objects; interactive objects use Mesh while backgrounds use 3D Gaussian Splatting, and a SimReady mode estimates physical attributes like mass and friction.
Applications span embodied AI, gaming, film and XR: partnerships with Dijie Robotics and Mouxianfei for robot simulation training, a July collaboration with Unity China, and film workflows combining 3D scenes with video models like Seedance 2.5 for multi-shot consistency. The team notes it currently supports rigid bodies only, physical properties are inferred rather than measured, and industry adoption is still in beta/validation stages.
Related event: Deemos Releases WorldGen: One Image Generates Fully Editable 3D Worlds(8 posts)→
More from Embodied
- Scoble claims ByteDance has a 120-gram AR device in labs that's blowing people away — Scobleizer · 2026-09-03
- Day 30 of U-BOT: fable 5.1 agent wrote a WiFi/BLE joystick controller — _Stocko_ · 2026-09-03
- Robot "fights back" after being pushed in viral test clip — cixliv · 2026-09-03
- Unity researcher voxelized his living room into billions of millimeter-scale voxels, stunned Stanford VR audience — Scobleizer · 2026-09-03
- Meta's Muse Spark jumps 1.1 to 1.3 in 55 days, photo-to-3D simulation at $0.60 — alexandr_wang · 2026-09-03
- Counterfactual debugging scales sim2real failure diagnosis to 1M steps in world models — sarahcat21 · 2026-09-03