Fish Audio launches Drama 3 preview, billing it as the most controllable TTS model ever

bdsqlsz · x · 2026-09-24

Fish Audio has released Drama 3 (preview), which it calls the most controllable TTS model ever. Users describe tone, pacing, and character in plain language instead of audio tags, can shift voice mid-sentence, generate multi-character scenes, or fix a single word without regenerating the whole clip. Commenters can request an API key to try it.

Original post →

More from Multimodal

Multimodal channel →