FixAnything: Refining 3D Renders into Photorealistic Videos via Video Generative Priors
orlitany · x · 2026-08-27
CMU researchers present FixAnything, a generalist video model that repairs rendering artifacts from any 3D representation—3DGS, NeRF, meshes, or sparse point clouds—into photorealistic, 3D-consistent videos. By repurposing a pretrained video diffusion model (Wan2.1) and using Flow-DPO with camera pose accuracy as a reward, the model achieves geometric consistency. Even very sparse point clouds serve as effective camera control with minimal LoRA finetuning.
More from Multimodal
- H3 Max generates 'Master Chief visits Seinfeld' in 6.6 seconds — chrisfirst · 2026-08-27
- Nvidia Releases 4-Step Versions of Cosmos3 Super T2I and I2V Models — q5sys · 2026-08-27
- ComfyUI and MiniMax Launch H3 Sync Sound Challenge — Comfy-Org · 2026-08-27
- Gen2Physics: Grounding 3D Meshes in Physics via Multi-View Material Decomposition — kwangmoo_yi · 2026-08-27
- Reddit user: Renting a GPU yourself is ~10x cheaper than third-party AI video generation sites — Forsaken-Low4467 · 2026-08-27
- Vercel AI Gateway Integrates Meta's Muse Image Model — evilrabbit_ · 2026-08-27