OrcaDub 1.0 launches speech-to-speech video dubbing with 4.83 MOS
qinzytech · x · 2026-07-28
OrcaDub 1.0 aims at production-grade speech-to-speech video dubbing
OrcaDub says its new model can localize videos end to end while preserving voice identity, emotion, prosody, timing, and lip sync.
Reported results
- 4.83/5 MOS
- 96.8% speaker similarity
- 92.7 COMET
- 3.1% WER
- 98.5% speaker attribution
What it offers
- Speech → speech dubbing
- Emotion and prosody preservation
- Idiom-aware localization
- Song translation
- Per-word forced alignment
- Background-music preservation
- Free web studio for post-editing
- OpenAI-compatible API
The service is available now with PAYG pricing at $0.60/minute and 10 free minutes.
More from Multimodal
- A LoKr test on Krea2 looks a lot like LoRA, with decent early results — Beautiful_Egg6188 · 2026-07-28
- Kinetic Abyss shows a 2D animation made with Stable Audio 3 and LTX 2.3 — Tadeo111 · 2026-07-28
- How to turn reference photos into a convincing impasto oil-paint workflow — Suspicious-Koala-570 · 2026-07-28
- A Midjourney style workflow uses GPT breakdowns to narrow hundreds of image references — beechinour · 2026-07-28
- Krea2 Detail Enhancer LoRA Removes AI Fog Blur for Crisp Images — linoy_tsaban · 2026-07-28
- Creating a 2000s Pop-Punk Music Video Using Runway and Suno — notiansans · 2026-07-28