AI Philosophy: Intrinsic interest in models without agenda
voooooogel · x · 2026-08-22
The post discusses the motivation behind AI philosophy, arguing that understanding how LLMs "grok" morality has terminal philosophical interest, independent of alignment relevance. It reflects on the linguistic community's dismissal of GPT-3 and the author's pivot to mechanistic interpretability.
More from AGI Musings
- AI Interaction Will Leap from Text to Mixed Reality — mrjonfinger · 2026-08-22
- DarkHaven alliance forming to fight AI doomers — beffjezos · 2026-08-22
- The irony of AI progress: computers may be the first to prove P=NP — EricBuess · 2026-08-22
- Philosopher to Discuss AI Ethics and Teleportation at WorldCon — eschwitz · 2026-08-22
- "Nobody has priced in the Singularity": a viral satirical thread imagines post-AGI daily life — louisvarge · 2026-08-22
- Survey Shows Young People Deeply Distrust AI CEOs — GaryMarcus · 2026-08-22