Nari Labs launches 50ms TTS endpoint, 10x cheaper than ElevenLabs on Qwen3-TTS

iamaliveix · x · 2026-09-11

Nari Labs launched a TTS inference service claiming a 50 ms time-to-first-audio, the fastest endpoint around, at $5 per 1M characters. Built on Qwen3-TTS 1.7B with their own inference engine, it's 5x faster than Cartesia and 10x cheaper than ElevenLabs, with expressive voices said to beat industry averages. Free for a limited time; the team frames open source as the future of multimodal AI, starting with speech.

Original post →

More from Infra

Infra channel →