omarchy-cluster runs the full 753B-param GLM-5.3 across four old Macs as one endpoint

natesiggard · x · 2026-10-08

Developer Joshua Warren released omarchy-cluster, an open-source project that lets networked machines run one local MLX model served as a single OpenAI-compatible endpoint. In the demo, a Mac Studio, three MacBook Pros and a Dell laptop on Omarchy jointly ran GLM-5.3 — 753B parameters, 216 GB of weights — as the full model, not a quantized one.

A reproducible recipe for giving old Macs a second life in local LLM inference.

Original post →

More from Infra

Infra channel →