Magnitude: open-source inference engine that self-tunes kernels, up to 2x faster than llama.cpp
paranoidray · reddit · 2026-10-01
Magnitude is an open-source local inference engine (in the vein of LM Studio / Unsloth Desktop) that compiles and tunes its kernels on your device for your exact hardware:
- Claims open models run up to 2x faster than llama.cpp thanks to the per-device optimization
- Works on Apple Silicon, NVIDIA, AMD, or CPU-only machines
Available on GitHub (magnitudedev/magnitude).
More from coding & agent
- Dev uses AI render projection to build Silent Hill-style Three.js kitchen with Astra — nptacek · 2026-10-01
- Tweet 'I need an app that...' and someone ships you an MVP in 30 minutes — thejasminejade · 2026-10-01
- 7-person robot startup Innate says Grok and Cursor let them do the work of 50 — Baconbrix · 2026-10-01
- Dev uses OpenAI's Dot to seamlessly monitor Claude Code benchmark on his machine — dkundel · 2026-10-01
- ChatGPT Sites Can Now Host MCP Servers, Auto-Deploying Plugins Across Platforms — OpenAIDevs · 2026-10-01
- Omnigent 0.16 ships with new file browser, admin controls and smoother onboarding — matei_zaharia · 2026-10-01