OpenBMB releases bilingual UltraData-RL dataset for RL training
openbmb · hf · 2026-09-08
OpenBMB's UltraData-RL-2609 dataset is trending on Hugging Face: a bilingual (English/Chinese) text-generation and QA dataset of 10K-100K rows under apache-2.0, tied to two arXiv papers and aimed at RL training data.
More from Research
- Mitra-v2: Synthetic-Data-Only 77M Tabular Model Matches 1.6B Rivals on TabArena — chaumian · 2026-09-08
- Sony AI's Hakken system turns 1.5M aging hypotheses into 2 confirmed gene discoveries — i_dg23 · 2026-09-08
- New worklog details building an async RL framework from scratch in JAX, from multi-actor systems to weight sync — yoshiyama_akira · 2026-09-08
- TASTE: A New Benchmark Testing If Models Can Predict AI Safety Researchers' Preferences — burny_tech · 2026-09-08
- EMNLP 2026 Paper: Training World Models for Behavior Consistency Cuts False Positives from 42.5% to 9.5% — 机器之心 · 2026-09-08
- Track4World: HKUST and Tencent ARC's feedforward model densely tracks every pixel in 3D — rsasaki0109 · 2026-09-08