audio.cpp 0.6 Released: Adds MiniMax-H3 Text-to-Audio, 5 New Model Families
Acceptable-Cycle4645 · reddit · 2026-08-17
audio.cpp 0.6 adds 5 new model families (dots.tts, NeuTTS-2e, MuScriptor, MiniMax-H3, SenseVoice-Small), totaling 49 families and 70+ variants. Highlights include native WebUI, MiniMax-H3 text-to-audio pipeline (TTS/voice clone/music gen), and MiniMax-Music3 preview. The H3 implementation can also produce video frames.
More from coding & agent
- WhatsApp Blocks New Device Linking? Dev Tries 4 Libraries, All Fail — Jason-Ping · 2026-08-17
- ComfyUI Mixed Mode for MiniMax H3: Multiple Generation Modes in One Timeline — Acceptable-Chest9695 · 2026-08-17
- MiniMax H3 Mixed Mode Deep Dive: Per-Segment Generation Modes — Acceptable-Chest9695 · 2026-08-17
- User observes ChatGPT quality degrades in new chats compared to long threads — DawniJones · 2026-08-17
- Red Hat: Building a production-grade operational layer for AI agents — blaizedsouza · 2026-08-17
- The Four Types of Agent Loops: Choosing the Right Structure for Your Task — blaizedsouza · 2026-08-17