Thinking Machines Unveils Inkling-Small Model

Thinking Machines Lab has officially released the Inkling-Small model. Built on a Mixture-of-Experts (MoE) architecture, this natively multimodal model features a total of 276B parameters while activating only 12B per token and supports a context window of up to 1M. Its weights are now fully open-sourced, with both the official team and the community highlighting its significantly lowered deployment barriers alongside outstanding performance.

已确认

为什么重要

2026-07-31 ~ 2026-07-31 · 22 related posts

Primary sources

5 near-duplicate retellings: simonguozirui · simonguozirui · ArtificialAnlys · simonguozirui · baseten