Open pretraining run matches Llama 3.2 1B on ARC-C at ~10% of the cost, author details 4 pitfalls

jon_durbin · x · 2026-10-03

jondurbin shared preliminary results and a post-mortem of the Kappa pretraining run:

Results: At the 576B-token checkpoint, the model beat Llama 3.2 1B (trained on 9T tokens) on ARC-C, OBQA, TQA and others, and nearly matched it on ARC-E, SciQ, etc. — at roughly 90% lower cost per token.

Issues encountered:

He calls it "a solid preliminary win" and plans fixes.

Related event: Open Pretraining Run Matches Llama 3.2 1B at One-Tenth the Cost(2 posts)→

Original post →

More from Infra

Infra channel →