Nari Labs Open-Sources Qwen3-TTS Engine, Beats ElevenLabs on Accuracy and Price
toebee · hn · 2026-09-15
Nari Labs (makers of Dia) launched Qwen3-TTS/ASR endpoints and open-sourced a custom inference engine for speech models. On Coval benchmarks, their TTS ranks #1 in accuracy (beating ElevenLabs and Cartesia) and is the cheapest endpoint; ASR has the lowest latency and #2 accuracy. Key insight: vLLM/SGLang fit multimodal inference poorly, so they built a specialized engine achieving sub-50ms latency at 10 RPS — outperforming even Alibaba's official endpoints. Next up: diarization, video and world-model inference.
More from coding & agent
- Cloudflare ships granular authz for Workers, its most requested feature — dinasaur_404 · 2026-09-15
- px0 adds millisecond full-text search, project-wide file open and auto-reload — arpit_bhayani · 2026-09-15
- Cloudflare adds granular authz to Workers with four roles for teammates and agents — dinasaur_404 · 2026-09-15
- Pick one coding agent harness, learn it deeply, and stop switching tools — iannuttall · 2026-09-15
- Hypothesis creator hiring London engineers for Hegel, open-source property-based testing — xuanalogue · 2026-09-15
- Adaption AI launches Invent API: training datasets from a prompt, no strings attached — sarahookr · 2026-09-15