LittleLearner: 88B-token grade-school-only pretraining shows LLM skills are elicited, not acquired

amplifiedamp · x · 2026-10-07

Researchers from MPI Tübingen, ELLIS and ETH Zürich built LittleCurriculum, an 88B-token corpus distilled from FineWeb-Edu via a five-stage filtering pipeline aligned to US Common Core K–5 standards, explicitly excluding anything taught above grade 5.

They trained 0.6B/1.3B/5B models from scratch on it, each paired with a matched unfiltered control (same architecture, tokens, recipe — only the corpus differs), creating a controlled sandbox with an interpretable knowledge boundary. A 5B model is chatable live in the browser.

Key finding: elicitation, not acquisition. Scaling, SFT+GRPO post-training, and in-context learning only amplify what the curriculum taught — they don't unlock out-of-curriculum knowledge. On mixed corpora it's thus hard to tell whether a new skill was learned or merely elicited. One reply proposes applying the same pedagogically-controlled approach to study consciousness.

Original post →

More from Research

Research channel →