Fast audio re-gen trick: downscale video latents to fix Turbo LoRA's bad audio
xyzdist · reddit · 2026-09-03
A practical fix for degraded audio from Turbo LoRA / low-step video generation: save latent + conditioning, reload with downscaled latent resolution (0.5x) since only audio matters, drop the LoRA and raise steps to 30+, then lock the video latent or set denoise 0.5 to keep audio aligned with visuals. Runs in about a minute; similar to the audio-refine custom node.
More from Multimodal
- Gemini's 88% video token cut landed on 3.7 Flash, not the 3.8 everyone's talking about — Servola-Journal · 2026-09-03
- Open-source ArtSmoker pipeline turns text prompts into textured, Blender-ready 3D models on your own AWS — niravdd · 2026-09-03
- SonicCaps: 15M-caption audio dataset improves CLAP retrieval and zero-shot classification — serrjoa · 2026-09-03
- Midjourney character sheet + Seedance long video recreate a 70s Giallo thriller — Kyrannio · 2026-09-03
- Catwoman character explorations generated with Midjourney — hewarsaber · 2026-09-03
- Fixing AI anime background inconsistency by pre-shooting a 360° orbit video with MiniMax H3 — Hailuo_AI · 2026-09-03