Geek Test: Running a 122B Parameter LLM on an Obsolete Laptop

_TheGreatDreamer_ · reddit · 2026-08-13

A hardcore developer successfully ran a massive 122 billion parameter Qwen model on an obsolete laptop, predating the LLM era, purely for testing purposes.

Although constrained by the hardware, the model took 5 minutes to load, 2 minutes to process the prompt, and 14 minutes to generate a response. However, this extreme test proves the powerful advancements in quantization and inference tools like llama.cpp—showing that running ultra-large models locally is no longer an absolute physical impossibility, even on extremely weak hardware.

Original post →

More from Infra

Infra channel →