Fish Audio launches Drama 3 preview, billing it as the most controllable TTS model ever
bdsqlsz · x · 2026-09-24
Fish Audio has released Drama 3 (preview), which it calls the most controllable TTS model ever. Users describe tone, pacing, and character in plain language instead of audio tags, can shift voice mid-sentence, generate multi-character scenes, or fix a single word without regenerating the whole clip. Commenters can request an API key to try it.
More from Multimodal
- Viral prompt turns your photo into a cinematic movie scene while keeping identity — Aiden_Tech_Ai · 2026-09-24
- Stop saying "edit my photo": 7 ChatGPT image-editing prompts that actually work — Aiden_Tech_Ai · 2026-09-24
- Skip ComfyUI: generating images straight from the terminal is underrated — JayoTree · 2026-09-24
- A 15-second one-take race car video prompt with frame-by-frame camera specs — umesh_ai · 2026-09-24
- AI UGC video pipeline: Hypit for strategy + HeyGen for digital humans, cheaper than video models — yihui_indie · 2026-09-24
- Redditor Shares Qwen 2.1 Character LoKr Training Recipe: Multi-Resolution Is Key — Any_Tea_3499 · 2026-09-24