Solar Open 2 tops Korean benchmarks with an 85.4 average score
keunwoochoi · x · 2026-07-26
- The post says Solar Open 2 performs very well on Korean tasks and points to appendix examples from regulatory documents.
- The attached benchmark screenshot shows Solar Open 2 posting the highest average across six models on the Korean suite, at 85.4, ahead of DeepSeek-V4-Flash, GPT-5.4 mini, and Claude Haiku 4.5.
- It also appears near the top on several Korean-language and professional-domain benchmarks such as CLICK, KBank-MMLU, KBL, and KMMU-Pro.
Related event: South Korea's Sovereign LLM Solar Open 2 Leads in Korean Benchmarks(2 posts)→
More from Models
- Fable 5 beats Opus 5 on open-ended coding tasks, says developer after 24 hours — johnlindquist · 2026-07-26
- Anthropic users say Opus 5 is best for complex agents, while Sonnet 5 fits routine coding — dr_cintas · 2026-07-26
- Reddit users question whether 5.6 Sol still has a Pro-only model tier — Massive_Sherbert_152 · 2026-07-26
- Frontier labs are racing to commoditize models as coding performance matters more — willccbb · 2026-07-26
- Google and NVIDIA leaders back open-weight models as essential to AI progress — omarsar0 · 2026-07-26
- Grok 4.5 posts Augment Code’s biggest week-over-week usage jump — XFreeze · 2026-07-26