Local Speech-to-Text Suite Released
tom_doerr · x · 2026-07-19
This is a local speech-to-text application named Transcription Suite: it features speaker diarization, long-audio and real-time transcription, cross-platform support, and an OpenAI-compatible AI assistant interface.
It is built on multi-backend ASR: Whisper, NVIDIA NeMo, VibeVoice-ASR, and whisper.cpp. It also supports GPU acceleration, AMD/Intel Vulkan, Apple Metal, and CPU modes, and is fully Dockerized for quick deployment.
More from Apps
- A market map tracks outpatient healthcare agentic AI across front, back and mid office — HealthcareAIGuy · 2026-07-21
- Synthesia launches Dubbing 2.0 with 130+ languages and lip-sync video translation — synthesiaIO · 2026-07-21
- User plans dozens of voice interviews with ChatGPT to build a book about themselves — mikesimmi · 2026-07-21
- Linear Launches Loops: Automate Workflows with Plain English Instructions — xiaohu · 2026-07-21
- Halliday opens priority access to G2 display AI glasses for meetings and daily use — SucceededMind · 2026-07-21
- A homework-photo app found the hard part is not OCR but date ambiguity and task splitting — Hayk_D · 2026-07-21