MiniMax M3 Released with 1M Context and SOTA Coding Benchmarks

MiniMax_AI · x · 2026-08-27

MiniMax has released the M3 model, available on SambaCloud and designed for long-horizon agents. It features a 1M-token context window and introduces MiniMax Sparse Attention (MSA), delivering over 9x faster prefill and 15x faster decoding than its predecessor.

On benchmarks, M3 achieved 59.0% on SWE-Bench Pro, 66.0% on Terminal-Bench 2.1, and 74.2% on MCP Atlas. In internal testing, it autonomously optimized a CUDA kernel for 24 hours, boosting hardware peak utilization from 7.6% to 71.3%. The model is natively multimodal (text, image, video), capable of understanding charts and controlling desktops.

Original post →

More from coding & agent

coding & agent channel →