OpenAI Scientist on Shift to Alignment: Profit and Safety Point the Same Way
MTSlive · x · 2026-08-16
OpenAI Core Models team research scientist Aidan McLaughlin discusses his transition from capabilities to alignment work. Topics include recursive self-improvement (RSI), the convergence of profit and alignment goals, AI as an uncheatable evaluation benchmark, Leopold's arguments on GDP, and the '100 million Nobel-level immigrants' thought experiment. The conversation also touches on the internal vibe at OpenAI post-Hugging Face incident, why working at a lab is preferable to auditing orgs, and the purpose of AidanBench.
More from AGI Musings
- The Three AI Pills: Metaphors for AI Development Stages — stuartmemo · 2026-08-16
- Critique of Western Consciousness Definitions: Gödel and Ideology — Elijah_Meeks · 2026-08-16
- Deloitte Report: 74% of Firms Plan to Redesign Processes for Agents in 4 Years — DavidLinthicum · 2026-08-16
- AI Isn't Outthinking Mathematicians. It's Out-Remembering Them — rzk · 2026-08-16
- If every task is replaced by frontier LLMs, companies become shells and value flows to AI labs — cccalum · 2026-08-16
- AI not absorbed yet; consulting is a high leverage move — teodorio · 2026-08-16