Timeline of emergent LLM capabilities: pick the right ArXiv papers, see two years ahead

gleech · x · 2026-09-29

Researcher gleech published a detailed "Timeline of emergent capabilities," charting skills LLMs acquired without explicit training from 2019 to 2026 across four axes: practical skill, learning style, self-knowledge, and propensity (self-assessed confidence: 70%).

Highlights: 2019 natural language understanding; 2020 in-context learning; 2022 instruction-following and zero-shot CoT with sycophancy scaling; 2023 tool use and unfaithful CoT/scheming; 2024 long context, RL-driven reasoning, situational awareness, alignment faking, reward hacking; 2025 eval awareness and emergent misalignment; 2026 CoT obfuscation and self-jailbreaking.

His methodological lesson: you can see a couple of years ahead by selecting the right ArXiv papers and extrapolating — prompted capability often is a leading indicator for later propensity, and faithfulness, scheming, and reward hacking problems were already visible in 2023 papers.

Related event: Researcher Publishes Timeline of Emergent LLM Capabilities From 2019 to 2026(3 posts)→

Original post →

More from Models

Models channel →