Fish Audio Makes S2.1 Pro Voice Model Free, Costs 1/6 of ElevenLabs
rohanpaul_ai · x · 2026-07-30
Fish Audio has made its S2.1 Pro voice cloning model free for a month. The model requires only 10-15 seconds of audio to clone a voice, supports 83 languages, and features a low latency of 90ms with word-level pronunciation and pause control.
Architecture & Ecosystem: It uses a Dual-AR design (a 4B-parameter model for semantics/emotion and a 400M-parameter model for acoustic details) balancing speed and expressiveness. Built on their open-source project Fish Speech, they continue to offer open-weight models for self-hosting. Pricing is roughly 1/6th of ElevenLabs.
Core Advantage: Tuned for real conversation, the model is designed to survive interruptions, corrections, laughter, and language switches.
Related event: Fish Audio Launches S2.1 Pro Voice Model and Raises $52M(11 posts)→
More from Models
- Google Launches Gemini Robotics 2: Single Model Enables Multi-Robot Collaboration — DynamicWebPaige · 2026-07-31
- OpenAI Cuts Terra and Luna Model Prices, Luna Down by 80% — bindureddy · 2026-07-31
- Google Launches Gemini Robotics ER 2 Embodied Reasoning Model — rseroter · 2026-07-31
- Inkling-Small Released: 276B Parameter MoE Model Matches Original Performance — ziqiao_ma · 2026-07-31
- OpenAI Launches GPT-5.6: Luna Costs 80% Less, Terra 20% Less — kiki-le-koala · 2026-07-31
- OpenAI Kills Price Advantage of Chinese Open-Weight Models, Threatening Market Takeover — arrakis_ai · 2026-07-31