Saudi HUMAIN launches free Voice model with Saudi dialect STT/TTS and developer SDK
aziz4ai · x · 2026-09-04
HUMAIN, Saudi Arabia's AI company, has launched HUMAIN Voice, a free-to-try (no signup) speech platform covering STT, TTS and an upcoming Voice Agent, with notably strong Saudi dialect quality.
Developer-side, it ships a unified SDK in TypeScript/Python (Go coming soon):
- File transcription with word-level timestamps, including the Arabic-English BayanArEn ASR model
- Realtime streaming transcription and live speaker diarization
- Streaming PCM TTS output
- Utilities like personal-data redaction and number/date formatting
Related event: Saudi HUMAIN Releases Voice Model with Arabic Dialect Support(2 posts)→
More from Multimodal
- Google releases Lyria 3.5 music model in AI Studio, Gemini API, and Gemini app — GeminiApp · 2026-09-05
- Microsoft launches MAI-Image-2.6-Flash: 2x faster than GPT-Image-2, 72% more GPU-efficient — mustafasuleyman · 2026-09-05
- Google ships Lyria 3.5 music model in Gemini with richer vocals and arrangements — GeminiApp · 2026-09-05
- Fei-Fei Li's World Labs unveils Atlas world model: 3 photos replace 300 for 3D capture — theworldlabs · 2026-09-05
- GPT Image 2 turnaround sheets + Seedance 2.5: a reproducible AI UGC ad workflow — socialwithaayan · 2026-09-05
- Topview and Wan 3.0 launch AI video challenge with $15,000 in prizes — azed_ai · 2026-09-04