ElevenLabs MCP adds voice, music, image and video generation to your AI assistant
aftahi_ai · x · 2026-09-24
- ElevenLabs launched voice, transcripts, dubs, music, sound effects, image and video generation inside its MCP server, letting assistants trigger full media workflows directly.
- User @Solethal shared a real save: a corrupted VO file that used to take three days to recover was fixed in 20 minutes with the new tools, rescuing a week of shooting.
More from Multimodal
- DeepMind ships Gemini 3.8 Flash TTS with native multi-speaker overlap and laughter — kastnerkyle · 2026-09-24
- Treating AI video as a production pipeline: building a reusable cinematic producer agent in CREAO — FellMentKE · 2026-09-24
- Open-source ComfyUI node restores lost template filters for Local, Partner and Credit workflows — linus74RN · 2026-09-24
- Gemini 3.8 Flash TTS demo nails the Osaka auntie accent — heiga_zen · 2026-09-24
- First/last-frame video workflow with GPT Image + LTX, author wants to move to local Qwen Image 2.1 — ART-ficial-Ignorance · 2026-09-24
- Gemini 3.8 TTS Hands-On: Emotion and Dialect Markers Enable Free Arabic Voiceover — aziz4ai · 2026-09-24