Fish Audio's Drama 3 lets you direct voice acting in plain language, shifts emotion mid-line
omarsar0 · x · 2026-10-08
Fish Audio launched Drama 3, a voice model you can direct in plain language — tone, emotion, pacing, and character — including shifting emotion mid-sentence without retakes. It's live as 'drama-3-preview' in the API. Blogger omarsar0 tested it and says it's the strongest directional control he's seen from a voice model, calling the human-like vocal control hard to believe.
Related event: Fish Audio launches Drama 3 voice model with mid-sentence emotion switching(2 posts)→
More from Multimodal
- Nano Banana 2.1 becomes Google's best image editing model across all 7 edit actions — ArtificialAnlys · 2026-10-09
- Nano Banana 2.1 Improves on All 9 Measured Image Capabilities, Biggest Gains in Layout and Anatomy — ArtificialAnlys · 2026-10-09
- Nano Banana 2.1 Pushes the Quality-Price Frontier: Three Ranks Higher at Half the Price — ArtificialAnlys · 2026-10-09
- Google's Nano Banana 2.1 Ranks #4 on Image Generation and Editing at Half Its Predecessor's Price — ArtificialAnlys · 2026-10-09
- Stanford's Level-of-Token Diffusion cuts image and video generation cost with multiresolution tokens — GordonWetzstein · 2026-10-09
- img2threejs: open-source image-to-3D rebuilds reference images as procedural Three.js models — tom_doerr · 2026-10-09