MiniMax voice models ignore gender cues in prompts, users report misassigned voices

DeltaWaffleSyrup · reddit · 2026-09-12

A Reddit user reports that MiniMax's voice generation models (FL2VA and REF2VA) largely ignore gender tagging in multi-character prompts, consistently assigning lower-pitched voices to male characters. Following official guidelines, toggling turbo LoRA, and raising steps to 20-30+ all failed. The only partial workaround is describing the female character as "a man in women's clothes", which breaks visual consistency. The author is asking the community for a real fix.

Original post →

More from Multimodal

Multimodal channel →