GUSH3R: Photorealistic Dynamic Human Scene Reconstruction from Monocular Video
janusch_patas · x · 2026-07-07
The paper introduces GUSH3R, tackling the novel problem of feed-forward, photorealistic, and renderable dynamic human scene reconstruction from monocular video. Core contributions include an architecture bridging human scene foundation models with photorealistic rendering, utilizing geometric priors and SMPL-X representations. It achieves competitive novel view synthesis quality compared to decomposition-based and optimization-based baselines while significantly boosting inference efficiency.
Related event: GUSH3R Enables Dynamic 3D Reconstruction from Monocular Video(2 posts)→
More from Multimodal
- LTX 2.3 LoRA demo changes a video’s camera angle — CQDSN · 2026-07-21
- Reddit user shares a surreal ChatGPT-generated poster — Creamy-Sundae-9991 · 2026-07-21
- A cinematic SEEDANCE 2 prompt turns an empty sunrise city into a memory-driven video — LudovicCreator · 2026-07-21
- Krea 2 Identity Edit transfers a karate pose from a line sketch without ControlNet — NatalieCrypto · 2026-07-21
- SenseTime unveils U1 Pro and open-sources a 50M-sample vision dataset at WAIC 2026 — 机器之心 · 2026-07-21
- Anatomy of Dynamic AI Images: Subject, Environment, and Camera — GPU_FieldNotes · 2026-07-21