Marigold V2: diffusion transformer depth estimation trained on one GPU in days
AntonObukhov1 · x · 2026-09-09
Marigold V2 (to appear at SIGGRAPH Asia 2026) is out. V1 post-trained an image generator into a depth estimator on a single GPU — the accessible research game; V2 upgrades to a diffusion transformer, producing very sharp edges in depth maps and normals, and is quite versatile.
Collaborator Matteo Poggi notes the whole thing trains in a few days on one GPU, no cluster needed, keeping the low-barrier, reproducible spirit alive.
Related event: Marigold V2: Single-GPU DiT Depth Estimator Hits Five Benchmarks(8 posts)→
More from Multimodal
- Fan-made foldable iPhone ad: one idea turned into a full AI commercial with Flova — Aiden_Tech_Ai · 2026-09-10
- Fixing AI Video's Scene-Rebuild Problem with a Blender-to-Seedance Pipeline — JaynitMakwana · 2026-09-10
- A no-drift AI video pipeline: Astra geometry to Blender to Dreamina rendering — JaynitMakwana · 2026-09-10
- GPT-6 Astra's 3D modeling falls far short of the demos in hands-on Blender testing — Next_Technology6361 · 2026-09-10
- Opus 5 makes a choir of singing faces, internet calls it "meditation music" — VoidStateKate · 2026-09-10
- Suno moves to licensed music training, but 'user data' from old models raises concerns — jordiponsdotme · 2026-09-10