Benchmark: MTPLX is the best engine to run Qwen3.8-27B on macOS

ex-arman68 · reddit · 2026-08-24

A rigorous benchmark of Qwen 3.8 27B on Mac Studio M2 Max compared multiple engines. MTPLX achieved the best balance of speed (20-24 tok/s) and score (91-93). llama.cpp + MTP was fastest in medium mode. The study concludes that 'xhigh' reasoning effort is worth the time cost if the engine is fast enough.

Original post →

More from Infra

Infra channel →