ESPnet's YODAS3 Speech Dataset with 1M+ Entries Trends on Hugging Face
espnet · hf · 2026-09-28
yodas3, a speech dataset from the ESPnet team, is trending on Hugging Face. It contains 1M–10M entries in parquet format under a CC-BY-3.0 license, covering ASR, audio-to-audio, TTS, and translation tasks across audio, text, and tabular modalities.
Related event: ESPnet Releases YODAS v3, Largest Open Speech Dataset at 1.1M Hours(2 posts)→
More from Research
- Cyber Index Alliance Launches With IBM, NVIDIA and Vercel to Standardize AI Cyber Defense Eval — ArtificialAnlys · 2026-09-28
- Artificial Analysis Details Cyber Index Methodology: Safety Refusals Score Zero, Tracked Separately — ArtificialAnlys · 2026-09-28
- Nissenbaum paper takes on privacy nihilism as AI inference erodes data-category frameworks — AllThingsApx · 2026-09-28
- Princeton researchers warn AI could slow science despite exploding paper output — AllThingsApx · 2026-09-28
- One-arm VR intervention during bimanual DAgger feels like an AI brain chip — neurosp1ke · 2026-09-28
- NeurIPS-rejected paper shows six agent dimensions to measure after benchmark saturation — random_walker · 2026-09-28