Tencent's AuK Tops Hugging Face Trending With Cloning, Editing and Separation
tencent · hf · 2026-09-11
Tencent's speech model AuK is trending on Hugging Face under the text-to-speech pipeline.
Its tags indicate a broad capability set: zero-shot voice cloning, speech generation and editing, speech enhancement, plus source separation and speech separation — making it an unusually all-in-one audio model.
More from Multimodal
- Testing Minimax H3 locally in ComfyUI: multi-prop reference images and a perfect 5s seamless anime loop — Time-Ad-7720 · 2026-09-11
- Curated collection of 153 GPT Astra prompts showcases one-shot 3D and game generation — TheMoonMidas · 2026-09-11
- DeepSeek V4.1 Flash multimodal: 45T image-text tokens and modality-level load balancing — nrehiew_ · 2026-09-11
- Is local AI video upscaling still broken? Hailuo 768p to 2K/4K求助 — Infinite-Emptiness · 2026-09-11
- Astra Can Compose SNES-Style Chiptunes, Vibe Coders Report — AIandDesign · 2026-09-11
- H3 Max video endpoint teaser blurs the line between real and AI footage — noahsolomon · 2026-09-11