Audio8 Open-Sources On-Device Audio Models From 0.1B to 3B for ASR and TTS
alexcovo_eth · x · 2026-09-08
Samuel Zeng's team has open-sourced multiple generations of on-device audio models via Audio8 over the past two months. The portfolio includes ASR models at 0.1B, 0.3B, 0.6B, and 3B, plus TTS models at 0.1B, 0.3B, and 0.6B — all designed for local inference on phones, PCs, and resource-constrained devices.
More from Multimodal
- Creator ships weekly AI series with InVideo agents, keeping creative control for free — LudovicCreator · 2026-09-08
- Grok-Generated Fake 2000s Seoul DV Home Video Goes Viral for Being Undetectable — SimplyAnnisa · 2026-09-08
- Sketch first, let ChatGPT paint over: a workflow where its image model shines — flowersslop · 2026-09-08
- Open-source GPT-6 Astra + Seedance 2.5 pipeline swaps characters in any video — matchaman11 · 2026-09-08
- MIT Multimodal AI Course 'How to AI (Almost) Anything' Videos Released Free — pliang279 · 2026-09-08
- WanGP fix unlocks longer viggle-animate videos beyond the 5-second limit — cocktailpeanut · 2026-09-08