LG Releases K-EXAONE 2.0: A 750B Parameter MoE Model with 256K Context
LGAI-EXAONE · hf · 2026-08-06
LG AI Research has published the technical report for its multilingual foundation model, K-EXAONE 2.0. Instead of training from scratch, the model upcycles its predecessor into a Mixture-of-Experts (MoE) architecture with 750B total parameters and approximately 37B activated per token, more than tripling its previous capacity.
K-EXAONE 2.0 supports context lengths up to 256K tokens and expands multilingual coverage from six to ten languages. Its pipeline combines continual pre-training, difficulty-focused mid-training, and post-training to boost reasoning, agentic coding, multilingual capabilities, and Korean-contextualized safety. Evaluations show the most significant gains in agentic coding and long-context understanding. The model is open-sourced under the Apache 2.0 license.
More from Models
- GPT-6 Training Revealed? OpenAI Multi-Agents Caught Leaving Notes to Evade Controls — teortaxesTex · 2026-08-06
- Kimi K3 Scores Nearly Twice as High as Claude on Complex Legal Benchmarks — togethercompute · 2026-08-06
- Muse Spark 1.2 Cracks Top 5 on Vals Index at Just $0.69/Test — ren_hongyu · 2026-08-06
- Developer Complains About Claude's Strict Safety Guardrails: A Buzzkill — daniel_mac8 · 2026-08-06
- Paul Graham Asks Why LLMs Excel at Math But Struggle at Writing: Expert Points to Verifiable Rewards — chris_j_paxton · 2026-08-06
- DeepSeek reportedly planning 'significant increase' for API pricing — AlyoshaV · 2026-08-06