FixAnything: Refining 3D Renders into Photorealistic Videos via Video Generative Priors

orlitany · x · 2026-08-27

CMU researchers present FixAnything, a generalist video model that repairs rendering artifacts from any 3D representation—3DGS, NeRF, meshes, or sparse point clouds—into photorealistic, 3D-consistent videos. By repurposing a pretrained video diffusion model (Wan2.1) and using Flow-DPO with camera pose accuracy as a reward, the model achieves geometric consistency. Even very sparse point clouds serve as effective camera control with minimal LoRA finetuning.

Related event: CMU's FixAnything Uses Video Priors to Fix Artifacts in Any 3D Representation(3 posts)→

Original post →

More from Multimodal

Multimodal channel →