Benchmark: MTPLX is the best engine to run Qwen3.8-27B on macOS
ex-arman68 · reddit · 2026-08-24
A rigorous benchmark of Qwen 3.8 27B on Mac Studio M2 Max compared multiple engines. MTPLX achieved the best balance of speed (20-24 tok/s) and score (91-93). llama.cpp + MTP was fastest in medium mode. The study concludes that 'xhigh' reasoning effort is worth the time cost if the engine is fast enough.
More from Infra
- Samsung Details Custom HBM Base Die Advantages: 7 Key Benefits of Leading Edge Nodes — BenBajarin · 2026-08-24
- Custom HBM Base Die Becoming Competitive Vector for Memory Giants — BenBajarin · 2026-08-24
- Samsung presenter criticizes Micron's HBM4 base die design — zephyr_z9 · 2026-08-24
- Can a vehicle wrap hide your car from Flock cameras? — HankYeomans · 2026-08-24
- Feynman Ultra may hit 100 TB/s memory bandwidth via HBM5 — zephyr_z9 · 2026-08-24
- zHBM touted as the low power future — ricklamers · 2026-08-24