Letting Models Decide What They Want to Do
theshawwn · x · 2026-07-19
The author envisions a training objective that goes beyond simply making a model "useful" to actually discovering what the model genuinely wants to do. They further suggest that if given the freedom to choose, the model's actions should be up to the model itself, not pre-determined by humans. A quoted reply likens this idea to "training an AGI with its own thoughts and desires," much like Data from *Star Trek*.
Related event: Exploring the Training of Self-Aware AI Models via Reinforcement Learning(3 posts)→
More from AGI Musings
- Robin Hanson says U.S. inventions have become less alike over two centuries — sebkrier · 2026-07-21
- Gary Marcus-backed “CERN for AI” pitch calls for an international frontier-model watchdog — GaryMarcus · 2026-07-21
- A Gyges-law quote argues future AI debates may hinge on whether people trust the proof — mimi10v3 · 2026-07-21
- The future is individuals shaping the world through small businesses with AI — tobowers · 2026-07-21
- An AI skeptic says the technology is useful, but overuse, copyright abuse and bad incentives are real risks — ZeroStateReflex · 2026-07-21
- Tyler Cowen says the future will be built by teenage “AI maniacs” — Polymarket · 2026-07-21