Tongyi Releases FlowTTS-GRPO Voice Cloning Model
udmrzn · x · 2026-07-14
Tongyi Voice AI introduced its FlowTTS-GRPO voice cloning model. Traditional voice cloning models often struggle to balance text accuracy, target timbre fidelity, and audio clarity/stability.
By applying Online Reinforcement Learning (Online RL) to flow-matching TTS, this new model successfully achieves simultaneous optimization across all three of these core dimensions.
More from Research
- Structural ensembles beat single predictions in TCR:pMHC generalization study — quaidmorris · 2026-07-22
- LLM leaderboards are now often measuring the harness too, Gary Marcus warns — GaryMarcus · 2026-07-22
- New paper defines self-state attacks, showing OS defenses leave four agent-memory cases indistinguishable — Justgototheeffinmoon · 2026-07-22
- Krea 2 users recommend a two-pass Clownshark sampler setup for sharper image details — listopalafoto · 2026-07-22
- Animation shows how an MLP’s first-layer weights change while learning MNIST — CatAstro_Piyush · 2026-07-22
- Project APE finds verifier reliability drops when papers contain multiple errors — soumitrashukla9 · 2026-07-22