Qwen 3.8 runs 151% faster on Apple Silicon via community challenge

gajesh · x · 2026-08-16

The MLX community launched a performance challenge for the Qwen 3.8 27B model on Apple Silicon. With community contributions, the model now runs 151% faster than the baseline (53.3 tok/s decode speed). The project explores acceleration limits for local LLMs using speculative decoding.

Original post →

More from Infra

Infra channel →