Building a production Greek-English speech recognizer: iterative training, filtering and ensembling
KIEFERSA · hf · 2026-09-15
A detailed engineering writeup of building a production-grade bilingual Greek-English speech recognition system.
- Iterative training, data filtering, ablation studies, and model ensembling.
- Strict quality gates throughout, achieving competitive benchmark results.
A complete industrial ASR engineering retrospective, valuable for anyone deploying multilingual speech recognition.
More from Research
- Microsoft's ESRL boosts MoE RL via expert-space exploration — MicrosoftResearch · 2026-09-15
- Grouped Value Attention shrinks KV cache by reconstructing keys on demand — Vishesh Tripathi · 2026-09-15
- Amazon's MInTRL uses sparse off-policy interventions to boost on-policy RL — amazon · 2026-09-15
- Stateless LLM failover preserves ~0% context; ContinuityBench proxy hits 99.20% CPR — its_vayishu · 2026-09-15
- Phillip Isola highlights a non-mainstream AI route: RL from scratch via ultra-fast simulators — AjdDavison · 2026-09-15
- Cutting AI verifier reading cost: top-50 retrieval kept just 2 of 8 minority evidence items — iMiguelmars · 2026-09-15