Bitter Lesson for RLMs: harnesses may drive generalization through decomposition
viksit · x · 2026-07-21
The post argues that the bitter lesson points away from domain-specific structure and toward general-purpose decomposition and recombination.
It draws an analogy to CPUs: once single-core clock speeds hit a heat wall, progress came from parallelizing workloads instead of just making one core faster. The author suggests LLMs may follow a similar path: as context windows and task horizons grow, scaling parameters will help, but big-O improvements in harnesses and decomposition may improve generalization even faster.
A quoted thread adds a more concrete claim: when training RLMs on tasks with shared structure that look different on the surface, the root model can learn the same trajectory for both. In that view, the harness—not the transformer itself—induces generalization by composing calls over structurally similar trajectories.
Related event: Research Suggests RLM Generalization is Driven by External Harness(10 posts)→
More from AGI Musings
- AI researcher on SkyNews flags concerns over inequality, power and criminal misuse — schwarzjn_ · 2026-09-11
- Researcher's SkyNews interview: deeply concerned about AI-driven inequality and power — schwarzjn_ · 2026-09-11
- VC compares AI doom rhetoric to pandemic-era fear messaging — StewartalsopIII · 2026-09-11
- Anthropic Insiders: Not Everyone at the Lab Believes in High p(doom) — anpaure · 2026-09-11
- Could 10k agents discover learning methods beyond backprop, or just tweak existing ones? — SeunghyunSEO7 · 2026-09-11
- AI companionship dissolves the friction real intimacy needs, warns long-form thread — YogeshMalik · 2026-09-11