JetBrains open-sources Mellum2.1 with scaled RL post-training for agentic coding

huggingface · x · 2026-10-08

JetBrains released Mellum2.1, a major update to the Mellum2 model open-sourced in June. The release comes from significantly scaling RL post-training to enable agentic coding workflows, achieving top agentic coding benchmark scores relative to its speed.

The 12B (MoE, 2.5B active) Thinking model weights are available on Hugging Face in both HF and GGUF formats, with MTP support coming in the following days.

Original post →

More from coding & agent

coding & agent channel →