Muse Glimmer 30B Local Test: 22 t/s on M5 Pro with 24GB RAM
curiousily_ · reddit · 2026-08-12
A developer tested Meta's new Muse Glimmer 30B model connected to the Hermes Agent for local and private AI coding tasks.
- Hardware & Performance: Achieved about 22 tokens/s on an M5 Pro device, utilizing around 24GB of memory including the drafter provided by Meta.
- Tool Calling: The model successfully executed correct tool calls and did useful work inside the Hermes Agent.
- Results: The resulting coding project actually worked, which was not the case when running the model with OpenCode.
Related event: Meta's Open-Sourced Muse Glimmer 30B Excels in On-Device Agentic Tasks(6 posts)→
More from coding & agent
- supastarter Updates Next.js SaaS Boilerplate for AI Coding Agents — jonathan_wilke · 2026-08-12
- Docker Sandboxes Are Reshaping the Agent Permission Model — krishnan · 2026-08-12
- Terraform Called 'Terrorism': Devs Debate Best IaC Tools for Modern Workflows — JasonBotterill · 2026-08-12
- Meta vs NVIDIA 30B Agent Models: Local Execution vs Cloud Routing — eyishazyer · 2026-08-12
- Stateless MCP 2.0 Spec Reignites Developer Interest — JeremyCMorgan · 2026-08-12
- Samsung Adopts Claude Code, Slashing Chip Verification Time from Months to Days — Wonderful_Buffalo_32 · 2026-08-12