Open-source Gepard TTS beats 24 closed APIs with 68ms latency on RTX 4090
ylankgz · reddit · 2026-08-27
Developers benchmarked the Apache 2.0 licensed Gepard TTS model against the Coval public leaderboard. Results show a median Time-To-First-Audio (TTFA) of 68.7ms on a single RTX 4090, outperforming all 24 closed-source APIs listed. The harness is open source, allowing users to benchmark their own models via vllm.
More from Infra
- ComfyUI INT6 Quantization Node Cuts Storage by 25% — BakaPotatoLord · 2026-08-27
- Chip sanctions backfire? SemiAnalysis says 100T free tokens per day — basedjensen · 2026-08-27
- Foresight Institute: Open science needs open compute as private monopoly hinders independent research — niloofar_mire · 2026-08-27
- Benchmark: Minimax H3 runs on 8GB VRAM with optimized attention — Zironic · 2026-08-27
- Google reveals 9,600-chip TPU 8t; OpenAI details 3-gen chip roadmap — SumitGup · 2026-08-27
- Optimizing inference on 4090: sub-10ms latency achieved — yacineMTB · 2026-08-27