M3 Ultra local benchmark: GLM5.2 and Qwen speeds limited

natesiggard · x · 2026-08-26

Testing GLM5.2 and Qwen 3.8 27B on an M3 Ultra with 512GB RAM, the user found GLM5.2 runs at 16 t/s (fine for background tasks) and Qwen 3.8 27B is capped under 30 t/s. Even if M5 doubles this speed, it won't match SOTA fast tokens for daily driving.

Related event: M3 Ultra Local LLM Tests Show Coding Still Out of Reach(2 posts)→

Original post →

More from Infra

Infra channel →