Why Papers Use Example Counts Over Token Counts
yanaiela · x · 2026-07-18
The author raises a specific question: why do many papers use "examples" rather than "tokens" when reporting post-training statistics?
This metric choice significantly affects how readers interpret training scale, data efficiency, and the comparability between different research efforts.
More from Research
- Talk at Geometry of ML 2026 shows AI finding and recommending resolutions to open math conjectures — wellecks · 2026-09-11
- Fast ViT shows strong ImageNet results; scaling runs needed next — ducha_aiki · 2026-09-11
- Loss Functions Are Scientific Assumptions: MSE Implies Gaussian Noise, Cross-Entropy Implies Bernoulli — bravo_abad · 2026-09-11
- SymKit MCP: 44 tools for AI agents to verify symbolic derivations — Foreign-Specific-604 · 2026-09-11
- Researchers: LLMs under pressure invent new languages unreadable to humans — mikeflache · 2026-09-11
- Mi-Ripple fixes ripple artifacts left by iterative AI image editing — Miyang-AI · 2026-09-11