Meta Releases MobileMoE: Sub-Billion Active Parameter On-Device Models
jacek2023 · reddit · 2026-08-25
- Family: MobileMoE is a family of on-device Mixture-of-Experts (MoE) language models with sub-billion active parameters, designed to push the quality-efficiency frontier for mobile LLMs.
- Scales: Available in three sizes (S/M/L) with 0.3B/0.5B/0.9B active parameters (1.3B/2.8B/5.3B total).
- Efficiency: INT4 weight footprints are <3 GB to fit in mobile DRAM.
- Variants: Each scale is released in three variants: Base (pre-training + mid-training), SFT (supervised fine-tuning), and QAT (quantization-aware training).
More from Models
- Chinese LLMs 4-5 Months Behind US; ECI 155 May Be Reliability Threshold — Jsevillamol · 2026-08-25
- Obscure board games as the best AGI eval: Fable far behind Opus 5 — paul_cal · 2026-08-25
- Codex vs Gemini vs Claude: Same Prompt, Wildly Different Results — thisiskp_ · 2026-08-25
- Nemotron 3.5 Lightning ranks top 4 open-weight models on Pinchbench agent tests — NVIDIAAI · 2026-08-25
- Agent Arena Pareto frontier: Claude and Kimi lead in cost-performance efficiency — arena · 2026-08-25
- AI Images Annoying in Technical Illustrations Due to Hallucinations — moultano · 2026-08-25