Cohere Releases Open-Source Arabic Speech Recognition Model
JayAlammar · x · 2026-07-07
Cohere has released an Arabic Transcribe model, claiming it to be the best open-source Arabic speech-to-text model available. It accommodates various Arabic dialects and has topped the Open Universal Arabic ASR Leaderboard. The model supports speech conversion between Arabic and English, as well as the recognition of English spoken with an Arabic accent, aiming to help developers and enterprises build robust voice experiences.
Related event: Cohere Releases Open-Source Arabic ASR Model, Topping Leaderboards(8 posts)→
More from Multimodal
- HeyGen adds a media-sourcing skill for coding agents with 75k images and 10k tracks — HeyGen · 2026-07-22
- Reddit user seeks ComfyUI NSFW text-to-image and image-to-video workflows under 20 GB VRAM — hobbyist2020 · 2026-07-22
- Krea 2 users recommend a two-pass Clownshark sampler setup for sharper image details — listopalafoto · 2026-07-22
- Gemini Omni Flash turns a boat cabin into a cave in Flow by Google — chrisfirst · 2026-07-22
- A simple workflow to turn a photo into an image prompt using Gemini, Grok, or GPT Image — harshitagu72595 · 2026-07-22
- A Reddit user proposes a consistency LoRA to keep anime and game scenes visually stable — ThirdWorldBoy21 · 2026-07-22