Kokoro-82M Enables Local Text-to-Speech
amos_gyamfi · x · 2026-07-14
Built on StyleTTS2, Kokoro-82M is an on-device text-to-speech model with 82 million parameters that supports 24kHz audio output.
The creator notes it can run on iOS 27 and macOS 27 via Core AI, providing links to Hugging Face and the Core AI Model Zoo.
More from Multimodal
- Seedance 2.0 demo turns ketchup on spaghetti in Rome into an AI reaction meme — azed_ai · 2026-07-21
- A reusable “Lunar Eclipse Dreamscape” prompt comes with multiple example renders — LudovicCreator · 2026-07-21
- Midjourney 8.2 preview shows a double-exposure prompt with strong style control — michaelrabone · 2026-07-21
- Travel MCP Server adds flight, hotel, weather and budget tools for agents — modelcontextprotocol · 2026-07-21
- Douyin Video Analysis MCP turns share links into structured video summaries — modelcontextprotocol · 2026-07-21
- Synthesia launches Dubbing 2.0 with 130+ languages and lip-sync video translation — synthesiaIO · 2026-07-21