Free Browser-Based Speech Recognition Tool Uses 3 Transformer Models for Diarization
doodlestein · x · 2026-08-13
A world-class multi-speaker speech recognition tool is gaining attention. It integrates three different Transformer models to provide not only high-quality ASR but also robust denoising and speaker diarization.
The tool is completely free, open-source, and runs entirely locally in the browser—meaning no audio data is uploaded, ensuring privacy. Tests show that while it may be slightly slower than the CLI version, the recognition quality is impeccable. It also supports exporting results as Markdown or self-contained HTML files for easy sharing.
Related event: Open-Source Rust Tool Enables Browser-Based Multi-Speaker ASR(2 posts)→
More from Apps
- Creating a Medieval Short Film Entirely with Adobe Firefly Boards — LudovicCreator · 2026-08-13
- Lottielab unveils Burp, an AI motion editor that made its own teaser without After Effects — Vjeux · 2026-08-13
- Stop Treating Claude Like an Intern: 8 Advanced Prompting Tips — hey_abusiddik · 2026-08-13
- Musk Retweets Praise for Grok Bot: Cross-Bot Messaging and Automations Work Perfectly — elonmusk · 2026-08-13
- Grok iOS App's GitHub Auth Flow Broken, User Reports — DanielLockyer · 2026-08-13
- HugoBlox: Build Academic Portfolios from Markdown with Auto-Imported Citations — tom_doerr · 2026-08-13