Running an Opus-level coding agent locally: MTPLX claims 126.5 TPS on Mac

julianharris · x · 2026-09-17

Julian Harris recommends MTPLX as the best way to run local AI coding agents, linking a deep-dive claiming a free, locally hosted Opus-level agent at nearly 2x Claude Opus speed. The referenced release, MTPLX V2.11.3, reports peak 126.5 TPS on a MacBook running a Qwen model, averaging 61 TPS at 100k context and 50 TPS at 200k context. After users complained about output quality vs LM Studio, the team ran thousands of sample comparisons, fixed 8 exactness bugs, and claims bit-exact outputs at any temperature. The author also vents about Anthropic's rate limits and censorship as a driver for going local.

Original post →

More from coding & agent

coding & agent channel →