MiniMax H3 Ignores Voice Volume Prompts; audio_reuse Offers a Partial Workaround
Dogluvr2905 · reddit · 2026-08-21
A user testing MiniMax H3 found that generated voices stay loud regardless of prompts like "quiet" or "distant mic," or even drastically reduced gain on the reference wav—H3 apparently reads waveform patterns, not volume. A workaround is audioreuse with a pre-lowered wav, but it replaces ALL audio in the clip, killing environmental sounds.
More from Multimodal
- Redditor Uses AI Video to Make a Literal Catfish: Cat Head, Fish Body — littleteckmonkey · 2026-08-21
- Xiaohongshu's FireRedTTS3 Unifies Speech Generation and Editing, Tops Benchmarks — 机器之心 · 2026-08-21
- Best repeatable AI image gen workflow after testing 99% approaches — EXM7777 · 2026-08-21
- Grok generates assets and edits migration video via chat — JoeJustice · 2026-08-21
- GitHub Trending: img2threejs Converts Images to Procedural 3D Code — anselm · 2026-08-21
- Midjourney's New Model Hailed as Extraordinary — doodlestein · 2026-08-21