Debate: Is RL the Devil? Pretraining-Only Routes vs OpenAI's Aggressive RL
jd_pressman · x · 2026-09-11
An X debate on training paradigms: @adrusi argues "RL is the devil" and capable AI is possible with just base models plus prostheses—harder to design but easier to reason about, with fewer welfare concerns, though race dynamics disfavor it. jdpressman quips that OpenAI is likely pan-frying its weights with the most cursed RL tricks Noam Brown can devise.
More from AGI Musings
- AI safety voices blur real risk with extinction talk, warns researcher: wrong fear is the worst fear — Dr_Atoosa · 2026-09-11
- AI safety insider: cheap extinction-risk talk is destroying the public's imagination of AI — Dr_Atoosa · 2026-09-11
- Hypothesis: ASI Has a Mathematical Incentive to Preserve Human Diversity — No_Cause_2731 · 2026-09-11
- New essay argues scaling laws originate in data structure, not architectures — gleech · 2026-09-11
- Seeking Optimism: Reddit Debates Whether Superintelligence Means Extinction — joshlove182 · 2026-09-11
- BlackHC: x-risk unlikely with current models, but rises sharply within a decade without changes — BlackHC · 2026-09-11