Puro-2B: an open recipe trains a Qwen2-1.5B-beating LLM on RTX 5090s for just $4.4K

IgorCarron · x · 2026-09-25

Puro-2B announces a fully open pretraining recipe for a 2B LLM that can be trained from scratch on consumer RTX 5090 GPUs:

Igor Carron ties it back to his 2021 essay "The $1,000 GPT-3": just as genome sequencing costs collapsed far faster than Moore's law once competition intensified, small-model pretraining costs are now undergoing a similar commoditization — progress doesn't come from one technology's steady march, but from a diversity of challengers.

Original post →

More from Infra

Infra channel →