MiniMax H3 voices sound too close and loud — user asks if an audio LoRA can fix the proximity problem
Dogluvr2905 · reddit · 2026-09-13
The author praises MiniMax H3 but flags a shared weakness: its generated voices (whether from reference audio or fully synthetic) are too loud and sound 'pasted in', as if the character is talking into a microphone a foot away, failing to sit acoustically in the scene. They've tried countless prompt combinations and feeding in low-volume reference audio — nothing works. The question: can a LoRA be trained to make H3's voices quieter and more 'distant' so they blend into the mix? A useful window into a real generation quirk: close-mic sound with no sense of distance or room ambience.
More from Multimodal
- Prompt template generates one travel scene in two styles: photoreal and watercolor side by side — nikola_mr64990 · 2026-09-13
- Mora 1 launches: AI-coded games with 3D generation and real-time video — fredodurand · 2026-09-13
- Designer asks: best generative tools for fast logo concept brainstorming? — RileyRalmuto · 2026-09-13
- World in World: training-free control of frozen video world models for re-camera and revisits — udmrzn · 2026-09-13
- Surreal AI-generated music video: The Jitterbug Jamboree — Hellish_NDE · 2026-09-13
- Open-source sound-and-vision app generates music and music videos for free — LawrenceOfTheLabia · 2026-09-13