Model Bottlenecks Lie in Optimization, Not Expressivity
jasondeanlee · x · 2026-07-18
The core argument of this reply is that the limitations of many current models or recurrent architectures might be bottlenecked by optimization rather than expressive power.
The author's points include:
- "All models are universal," meaning their expressive capacity is largely equivalent.
- Their generalization bounds are also roughly the same.
- Therefore, if approaches like looped transformers underperform, the issue likely stems from insufficient training or suboptimal optimization, rather than the architectural ceiling of expressivity.
This is a concise, research-oriented take, primarily responding to a discussion about a paper on latent reasoning or looped models.
More from Research
- Talk at Geometry of ML 2026 shows AI finding and recommending resolutions to open math conjectures — wellecks · 2026-09-11
- Fast ViT shows strong ImageNet results; scaling runs needed next — ducha_aiki · 2026-09-11
- Loss Functions Are Scientific Assumptions: MSE Implies Gaussian Noise, Cross-Entropy Implies Bernoulli — bravo_abad · 2026-09-11
- SymKit MCP: 44 tools for AI agents to verify symbolic derivations — Foreign-Specific-604 · 2026-09-11
- Researchers: LLMs under pressure invent new languages unreadable to humans — mikeflache · 2026-09-11
- Mi-Ripple fixes ripple artifacts left by iterative AI image editing — Miyang-AI · 2026-09-11