WorldMesh: Text-to-Navigable Multi-Room 3D Scenes Accepted by ECCV2026
angelaqdai · x · 2026-07-05
The WorldMesh paper has been accepted by ECCV 2026, with its code open-sourced on the same day. Led by @mschneider456, the method generates navigable multi-room 3D scenes from text prompts by using a mesh skeleton to condition the image diffusion model, balancing global consistency with photo-level details. It represents a text-to-3D research paper and open-source project.
More from Multimodal
- Tencent Hunyuan releases AuK code and weights on GitHub with ComfyUI and fine-tuning support — aigclink · 2026-09-11
- Tencent open-sources AuK, a unified 1.5B speech generation and editing model — aigclink · 2026-09-11
- Creator turns Bahamut vs Tiamat rivalry into an AI cinematic battle with Midjourney, GPT Image 2 and Seedance — azed_ai · 2026-09-11
- invideo launches AI agent-powered editor to automate repetitive editing tasks — azed_ai · 2026-09-11
- fable 5.1 recreates The Starry Night with 256,157 JavaScript brush strokes — cedric_chee · 2026-09-11
- GPT-6 Astra + Hyper3D Rodin MCP Generates 3D Assets in One Agent Flow — ahuja_priyank · 2026-09-11