Redditor runs gpt-oss-120b across a phone, three Macs and two Windows PCs

ANR2ME · reddit · 2026-09-29

Per wccftech, Reddit user "MedicineBlogscanner" managed to run the 4-bit quantized gpt-oss-120b — a model that officially requires at least 60GB of memory — on a wild distributed setup:

By pooling memory and VRAM across devices into a distributed inference setup, the user bypassed single-machine memory limits, showing consumer hardware can run 120B-scale open models through unconventional means.

Original post →

More from Infra

Infra channel →