DeepSeek signals: data quality ROI now beats novel post-training algorithms
A_K_Nain · x · 2026-09-11
A notable signal: DeepSeek — usually first to pioneer novel algorithms and architectures — says the ROI of improving data quality now far exceeds working on novel post-training algorithms.
The commenter argues this has been true for non-lab practitioners for a while: put 80% of your post-training effort into data —
- Hire experts to dig through RL tasks
- Manually sift rollouts and SFT data, remove suspicious samples, verify all tasks are solvable
- Ensure diversity in both difficulty and category
Related event: DeepSeek says better data now beats novel post-training algorithms(2 posts)→
More from Models
- Astra pauses $200 Pro subscriptions as demand hits unprecedented levels — emax · 2026-09-11
- inclusionAI's Open-Source Ling-3.0-flash-VL Multimodal Model Trends on Hugging Face — inclusionAI · 2026-09-11
- Tiny KV Cache via Shared Global KV Plus Per-Layer SWA? New Architecture Speculation — stochasticchasm · 2026-09-11
- DeepSeek v4.1 Flash Tested Across 8 Coding Harnesses: Performs Best in Minimal Setups — mariofilhoml · 2026-09-11
- Would ChatGPT Plus users accept a 24-hour usage limit instead of weekly caps? — SuaveSteve · 2026-09-11
- Leaked Qwen next-gen model shows record n-gram params, first two layers SWA-only — stochasticchasm · 2026-09-11