Running Qwen 2.5 on Mac Studio: Performance Tests and Agent Struggles

Over_Technology_1764 · reddit · 2026-08-16

A user tested Qwen 2.5 models (referencing 27B and 72B, despite a typo in the original text) on a Mac Studio with 64GB RAM. Using the community MLX 4-bit quantization yielded a speed of about 16 tps. When attempting to create a Pokémon game with Pi Agent, the model got stuck in a thinking loop and failed to progress with complex tasks. After unsuccessful optimization attempts with Hermes Agent, the user reverted to an older model and is seeking advice from others with similar setups.

Related event: Running Local Agent Models on Mac: Memory Is the Deciding Factor(2 posts)→

Original post →

More from Infra

Infra channel →