Nari Labs launches 50ms TTS endpoint, 10x cheaper than ElevenLabs on Qwen3-TTS
iamaliveix · x · 2026-09-11
Nari Labs launched a TTS inference service claiming a 50 ms time-to-first-audio, the fastest endpoint around, at $5 per 1M characters. Built on Qwen3-TTS 1.7B with their own inference engine, it's 5x faster than Cartesia and 10x cheaper than ElevenLabs, with expressive voices said to beat industry averages. Free for a limited time; the team frames open source as the future of multimodal AI, starting with speech.
More from Infra
- SpaceX CFO: vertical integration is core, Starship paves way for orbital compute — elonmusk · 2026-09-11
- Bezos: Power Supply Chain Bottleneck Forces AI Labs to Slow Development Pace — beffjezos · 2026-09-11
- vLLM upgrade guide: KV offloading, queue admission control, 33.6% Blackwell latency cut — vllm_project · 2026-09-11
- vLLM v0.29.0 cuts Blackwell E2E latency 33.6%, with 6.6-7.6x kernel speedups for Kimi-K3 — vllm_project · 2026-09-11
- One cheeseburger emits as much CO2 as 63,000 Gemini text prompts, math shows — recallingmemories · 2026-09-11
- Google signs deal to buy half the electricity of a nuclear power plant — lukaspetersson · 2026-09-11