NVIDIA's Kimodo: Open-source text-to-3D-motion model runs on RTX 3090
maier_ak · x · 2026-08-31
NVIDIA released Kimodo, an open-source diffusion model that generates full-body 3D human or humanoid-robot motion from text prompts while honoring keyframe, foot-target, or path constraints. The 282M-parameter model achieves an average joint-position error of 3.2 cm, 71.9% R-precision, and FID 1.85 on the 700-hour Bones Rigplay set. It runs on a single RTX 3090 or even GPUs with <3 GB VRAM.
Related event: NVIDIA Open-Sources Kimodo for Text-to-3D Human Motion Generation(4 posts)→
More from Embodied
- Tesla FSD Anticipates Swerve to Avoid Debris: User Story — surmenok · 2026-08-31
- Robot That Puts Your Shoes Away When You Get Home Launches in September — chris_j_paxton · 2026-08-31
- AI-powered intelligent plastic ducks hint at a future of droid toys — VoidStateKate · 2026-08-31
- NVIDIA's Kimodo Turns Text Into Full-Body 3D Human and Robot Motion — maier_ak · 2026-08-31
- Microducks robot sells over $2.5M in first 24 hours — _akhaliq · 2026-08-31
- Zhongqing's Shift from Hardware to Embodied AI Brains — 量子位 · 2026-08-31