Reliquary-4B: A 4B math & code model trained via decentralized RL with community rollouts

const_reborn · x · 2026-09-22

The Reliquary team released Reliquary-4B, a 4B model focused on math and code, trained with reinforcement learning in a decentralized way: anyone can join the network and contribute rollouts. Independent miners choose prompts and generate rollouts, while the protocol verifies them and trains the model.

The authors frame it as the start of an era of decentralized post-trained models that beat frontier models in their class. Model weights and the research behind it are public.

Original post →

More from Models

Models channel →