Inkling-Small Released: 276B Parameter MoE Model Matches Original Performance
ziqiao_ma · x · 2026-07-31
Thinky Machines has released Inkling-Small, a new open-weights MoE (Mixture of Experts) model. Featuring 276B total parameters and 12B active parameters, it achieves comparable performance to the original Inkling at a quarter of the size.
The full weights are available, allowing users to fine-tune the model on Tinker or chat with it in text, image, and audio on the Tinker Playground.
More from Models
- OpenAI Slashes GPT-5.6 API Prices and Boosts Speed — paw_lean · 2026-07-31
- Tinky Machines Releases Inkling-Small: 276B Parameters with Open Weights — simonguozirui · 2026-07-31
- Using "---" Confuses LLMs and Breaks Role Context — davidad · 2026-07-31
- GPT-5.6 Luna is Cheaper, Smarter, and Faster than Gemini 3.6 Flash — Angaisb_ · 2026-07-31
- LLM Inference Costs Plunge: Token Prices Drop to 1/13th in Four Months — charliermarsh · 2026-07-31
- vLLM Releases Inkling-Small Deployment Guide: Runs on Minimum 180GB VRAM — vllm_project · 2026-07-31