AI Optimism essay: AI is easier to control than human labor — a technical case for alignment
QuintinPope5 · x · 2026-09-09
Quintin Pope points to an essay, "AI is easy to control," building a technical case that alignment is feasible:
- More controllable than labor: AI is profitable precisely because its personality and conduct can be controlled with far finer precision than any human employee.
- The tooling: GPT-4 and Claude undergo supervised fine-tuning that directly optimizes neural circuitry to say specified things in specified contexts; DPO and RLHF holistically shape model behavior; curated pretraining data controls all of an AI's formative experiences.
- Economics: because AIs are cheaply copyable programs run in parallel, it's viable to invest millions or billions training a "single" artificial employee — impossible or unethical with humans.
- Conclusion: extinction-level AI takeover is implausible; even if future AI outpaces direct supervision, instilling human values (alignment) will be easy.
Related event: Quintin Pope Reiterates AI Is Easy to Control and Alignment Is Solvable(2 posts)→
More from AGI Musings
- OpenAI reportedly spent $15M in tokens on Navier-Stokes proof, and it was worth it — bindureddy · 2026-09-09
- On why evolution produced consciousness: intelligence as the plausible intermediary — jessi_cata · 2026-09-09
- Neal Stephenson has written by fountain pen for 25 years — Baroque Cycle manuscript stood 42 inches tall — ashishkr9311 · 2026-09-09
- Jeff Ladish clarifies: he means intellectual siloing, not just time on Twitter — JeffLadish · 2026-09-09
- Jeff Ladish: researchers building superintelligence being intellectually siloed is extremely risky — JeffLadish · 2026-09-09
- Why aren't you part of a system optimizing for subjective experience? An AI-consciousness argument thread — jessi_cata · 2026-09-09