Marigold V2 recipe: 4-bit quantized DiT + rank-128 QLoRA, fine-tuned on one 32GB GPU
AntonObukhov1 · x · 2026-09-09
Marigold V2 keeps unconventional choices that work. V1 pushed depth maps through an image VAE — no business working, but it did. V2 pushes ground-truth depth through DINOv3 and aligns DiT features to it (iREPA-depth). Recipe: take a large pretrained DiT (Qwen-Image-Edit-2509), quantize to 4-bit, add a rank-128 QLoRA, fine-tune on a single 32GB consumer GPU — sensible depth after a few hours, done in a few days, no 80GB cards or multi-node setups. Single-step inference, no OOM at 2K resolution.
Related event: Marigold V2 Released: Single-GPU Fine-tuned DiT for Depth Estimation(5 posts)→
More from Infra
- Fab2 raises $500M Series A at $3.7B valuation to scale chip fab business — Sethwinterroth · 2026-09-09
- Zach Dell on Base Power: meeting AI's energy demands, batteries and vertical integration — espricewright · 2026-09-09
- Smartphone brands raise prices in India as DRAM shortage seen lasting until late 2027 — saibharadwaj · 2026-09-09
- Inception ships Mercury 2.5: most capable diffusion LLM at 1,107 tokens/sec, 260K context — StefanoErmon · 2026-09-09
- Intel shows a 27x speedup in matrix multiplication by just swapping two loops, both O(n³) — jedisct1 · 2026-09-09
- AWS benchmarks G7 Blackwell vs G5/G6 for 30B MoE inference on SageMaker — AWS ML Blog · 2026-09-09