Trending on HF: 10M+ Multilingual LLM Distillation Dataset
r0b0tlab · hf · 2026-08-06
The qwen3.8-max-glm5.2-kimi-k3-distillation dataset published by r0b0tlab is trending on Hugging Face.
- Scale: Contains between 10M and 100M samples.
- Format & License: Available in Parquet format, licensed under 'other'.
- Use Case: Primarily targeted at text-generation tasks, supporting multiple languages including English, Chinese, Spanish, French, German, and Japanese.
More from Models
- LMSYS Launches Factuality Leaderboard to Rank AI Hallucinations — jfiance · 2026-08-06
- Hands-on with Ling-3.0-flash: free trial experience on AI/ML API — rohanpaul_ai · 2026-08-06
- Three Platforms Offer DeepSeek V4 Flash at One-Sixth of Official Pricing — oran_ge · 2026-08-06
- OpenSuperintelligence Lab Drops Neurosymbolic AGI System — xeophon · 2026-08-06
- How 19MB Vision Distillation Model Beats AI Giants in Specific Tasks — richdotca · 2026-08-06
- Hands-on: Meta's Muse Model Delivers Instant Replies on WhatsApp — DevDminGod · 2026-08-06