A 90M conversational LLM now runs on the 2004 Sony PSP at 0.5 tokens/sec

liright · reddit · 2026-09-05

A developer got a 90M-parameter conversational LLM running on the 2004-era Sony PSP, with the project open-sourced on GitHub. It runs at roughly 0.5-0.6 tokens per second — a reply takes 1-3 minutes — and 90M appears to be the practical ceiling for the hardware. The model can crank out bad poems, short stories and non-functional code, occasionally answers trivia correctly, and otherwise hallucinates wildly. Not useful, but delightfully extreme local AI.

Original post →

More from Fun

Fun channel →