Qwen3.8-27B runs at 63 tok/s on Mac Studio

remilouf · x · 2026-08-17

A user shared a benchmark of running the Qwen3.8-27B model locally on a Mac Studio, achieving an inference speed of 63 tokens per second, noting that it is very usable.

Original post →

More from Infra

Infra channel →