Supertonic: Open-Source Local TTS Released
JafarNajafov · x · 2026-07-14
Supertonic is a text-to-speech model that runs entirely on local devices, requiring no cloud, API keys, or per-character billing. The author states it is 100% open-source under the MIT license and has around 2700 GitHub stars.
The reported performance metrics are highly aggressive: it achieves 167x real-time speed on an M4 Pro with only 66 million parameters. In speed comparisons, Supertonic hits 1263 characters/second, compared to 287 for ElevenLabs Flash and 55 for OpenAI TTS-1. It can even run on a Raspberry Pi or offline e-readers.
Functionally, it handles currencies, dates, phone numbers, and technical units well without preprocessing. It supports 11 platforms and 5 languages, offering a Chrome extension to convert webpages to audio in under a second. The author concludes that the cloud API-dominated TTS market may soon be disrupted by such on-device models.
More from Infra
- How to build a PostgreSQL-backed semantic search pipeline with pgvector and Ollama — KhuyenTran16 · 2026-07-21
- NeurIPS 2026 workshop calls papers on on-device intelligence — YiMaTweets · 2026-07-21
- Milled from Solid Aluminum: AI Rig Multi-GPU Case for Local Compute — dee_hw · 2026-07-21
- FutureCaribbean’s Buildathon offers $50K, H200 compute, and an NYSE pitch — HeyAmit_ · 2026-07-21
- A new series tests which data-science workflows can run on GPUs today — pandeyparul · 2026-07-21
- Former AWS operator says Bedrock margins can beat SageMaker as agentic AI lifts CPU demand — RihardJarc · 2026-07-21