audio.cpp: pure C++ ggml engine runs TTS, ASR, voice conversion locally, no Python
solyarisoftware · x · 2026-09-10
audio.cpp is an open-source, pure C++ audio inference framework built on ggml, aimed at eliminating Python dependency hell for local audio models.
- Supports TTS, STT, VAD, voice conversion, and music generation
- Already integrates Qwen3-TTS, Qwen3-ASR, and more
- Cross-platform: Windows/Linux/macOS on NVIDIA, AMD, Apple Silicon, and CPU-only
- Ships with a WebUI, model comparison view, and unified API — handy for local voice assistants and dubbing tools
- 2.5k GitHub stars; CUDA backend is currently the optimization focus
More from Infra
- Hyperscalers could factor RSA-1024 for about $30M per number, analysis claims — rickasaurus · 2026-09-11
- Qualcomm's Next Hexagon NPU: 50% More Shared Memory, 30B MoE Models on a Phone — ryanshrout · 2026-09-11
- Vercel Cut CDN P99 Metadata Lookup Latency by 91% Across 80M Route Decisions/sec — cramforce · 2026-09-11
- OreoLook: three-layer caching for low-latency LLM web search on commodity CPUs — pollinations · 2026-09-11
- Analyst: NAND's DRAM moment — DeepSeek V4.1 Flash keeps 26% of params on SSD — casper_hansen_ · 2026-09-11
- NVIDIA BioNeMo kernels hit 7.1x speedup on triangle attention in Terray's drug discovery benchmarks — AllThingsApx · 2026-09-11