Diamond 1.0 Speech Restoration Model Goes Open Source
bdsqlsz · x · 2026-07-17
The open-source speech restoration model Diamond 1.0 has been released, aiming to restore degraded audio to studio-quality 44.1 kHz speech.
- Trained from scratch without using any pre-trained backbone model
- Has 1.666 billion parameters and required only 63 GPU hours for training
- Ranks 2nd out of 6 models on DNSMOS
- Reportedly outperforms models trained on four times the amount of data
- Released under the Apache 2.0 license and ready for trial
More from Multimodal
- Seedance 2.0 demo turns ketchup on spaghetti in Rome into an AI reaction meme — azed_ai · 2026-07-21
- A reusable “Lunar Eclipse Dreamscape” prompt comes with multiple example renders — LudovicCreator · 2026-07-21
- Midjourney 8.2 preview shows a double-exposure prompt with strong style control — michaelrabone · 2026-07-21
- Travel MCP Server adds flight, hotel, weather and budget tools for agents — modelcontextprotocol · 2026-07-21
- Douyin Video Analysis MCP turns share links into structured video summaries — modelcontextprotocol · 2026-07-21
- Synthesia launches Dubbing 2.0 with 130+ languages and lip-sync video translation — synthesiaIO · 2026-07-21