Qwen on M2 Ultra: latest oMLX update brings substantial local inference speedup

Thrumpwart · reddit · 2026-09-13

A Reddit user shares an update on running Qwen 3.8 Flash locally on an M2 Ultra: the latest oMLX release introduces a substantial speedup, with a benchmark screenshot attached. A notable signal for Mac-based local inference performance.

Original post →

More from Infra

Infra channel →