IFM releases K2-Horizon-MoVA-36B-A4B: 4B active params with 512K context
jacek2023 · reddit · 2026-09-03
IFM has released the final GGUF checkpoint of K2-Horizon-MoVA-36B-A4B, a sparse MoE model with Mixture-of-Values (MoVA) attention that stores 36B parameters but activates only 4B per token.
- Outscores open-weight dense (30B) and MoE models up to 15x its size on agentic and reasoning benchmarks, and competes with closed frontier models.
- Native 524,288-token context from midtraining onward.
- Fully open: training data/recipe, training code, and intermediate checkpoints will be released to study capability growth across training.
More from Models
- Gemini has improved exponentially, Grok slower and often unavailable: Damodaran — pdamodaran · 2026-09-03
- Hacker upgrades stolen account to $200 ChatGPT Pro with no fraud checks from OpenAI or bank — CucharitaDePalo · 2026-09-03
- Muse Spark 1.3 Update Gains Attention, Users Urged Not to Miss — D3VAUX · 2026-09-03
- 64GB Mac users debate best sub-40B MoE: is Qwen-3.8-35B worth the wait? — chibop1 · 2026-09-03
- xAI's mystery model 'Sol' spotted stealth-testing at 750 tps, rumored to drop soon — koltregaskes · 2026-09-03
- Microsoft's MAI-Transcribe-2 hits 2.0% WER at 411x real time for $1.67 per 1,000 minutes — ArtificialAnlys · 2026-09-03