Open Pretraining Run Matches Llama 3.2 1B at One-Tenth the Cost

Developer jondurbin shared early results from the Kappa pretraining run, matching Llama 3.2 1B benchmarks at roughly one-tenth the cost, with inference code achieving 23k tokens/sec decoding on a single RTX 5090.

2026-10-03 ~ 2026-10-03 · 2 related posts