SupraLabs releases a 5M-row reasoning corpus for tiny-model fine-tuning
LH-Tech_AI · reddit · 2026-07-24
SupraLabs released reasoning-corpus-4K-5M-v1, a 5-million-row dataset aimed at training tiny SLMs to reason.
- Each sample includes repoid, token length, user prompt, thoughttrace, assistant answer, and a ChatML-formatted record.
- All examples are capped at 5k sequence length to make them suitable for SFT and fine-tuning.
- The dataset is published on Hugging Face and the post invites feedback, follows, and reuse.
- The release is positioned as a large reasoning corpus for small-model training rather than a benchmark or model update.
More from Research
- Meta’s GAMUT benchmark scores long answers on missing facts, and the best model gets 58.7% — rohanpaul_ai · 2026-07-24
- Masked Visual Actions turns 15 hours of robot video into a zero-shot world model — jbhuang0604 · 2026-07-24
- New subnet design lets miners compete on training data instead of weights — const_reborn · 2026-07-24
- Cohere Labs opens a 48-hour model challenge on language learning and reasoning — Cohere_Labs · 2026-07-24
- A factor-model paper shows panel causal inference without parallel trends — PtrPomorski · 2026-07-24
- Ordinary Least Squares regression, explained with a simple fit chart — mdancho84 · 2026-07-24