Raschka would spend a $100M LLM budget entirely on post-training

Asked how he'd build a SOTA LLM with $100M, Sebastian Raschka said he'd skip pretraining and spend it all on post-training an existing model. A reply claimed Jev reportedly used 100% synthetic data with no pretraining, yet still generalizes impressively.

2026-09-20 ~ 2026-09-20 · 3 related posts