Solar Open 2 grows to 250B parameters while cutting total training tokens
keunwoochoi · x · 2026-07-26
- Solar Open 2 is described as a 250B sovereign LLM from South Korea, and the author says it is 2.5× larger than Solar Open 1 while using fewer total training tokens.
- The post highlights a four-stage post-training recipe: supervised fine-tuning, multi-domain RL, specialist training, and multi-teacher on-policy distillation.
- The paper also claims selective weight transfer from Solar Open 1, a 10T-token general pre-training stage, a 1T intensive stage, and Korean officework data curation for domain-specific performance.
Related event: South Korea's Sovereign LLM Solar Open 2 Leads in Korean Benchmarks(2 posts)→
More from Models
- Reddit users question whether 5.6 Sol still has a Pro-only model tier — Massive_Sherbert_152 · 2026-07-26
- Frontier labs are racing to commoditize models as coding performance matters more — willccbb · 2026-07-26
- Google and NVIDIA leaders back open-weight models as essential to AI progress — omarsar0 · 2026-07-26
- Grok 4.5 posts Augment Code’s biggest week-over-week usage jump — XFreeze · 2026-07-26
- Solar Open 2 gets quantized models and wider API access soon — keunwoochoi · 2026-07-26
- Solar Open 2’s pretraining trims 20T tokens to 10T and stretches context to 1M — keunwoochoi · 2026-07-26