Compute shift to post-training narrows open-closed model gap

A widely discussed essay argues that the compute shift from pretraining to post-training explains the closing gap between open-weight and proprietary models, since post-training advances are far easier to distill and harder to detect.

2026-09-15 ~ 2026-09-15 · 3 related posts

1 near-duplicate retellings: maksym_andr