Robert Long's open questions on introspection in empirical AI welfare
rgblong · x · 2026-10-02
Robert Long promotes his Substack essay listing in-the-weeds open questions in empirical AI welfare, focused on model introspection and skepticism toward Anthropic's 'functional emotions' framing. Same article as the previous post; see that entry for details.
Related event: AI Welfare Researcher Questions Anthropic's 'Functional Emotions'(2 posts)→
More from Safety
- Bipartisan AI Agent Accountability Act would hold developers liable for autonomous agent attacks — Miles_Brundage · 2026-10-02
- OpenSwitchboard: open-source MCP server gates agent commitments behind human presses — EnvironmentalRice348 · 2026-10-02
- Buyers now fill out export control declarations when purchasing RTX 5090s in stores — blelbach · 2026-10-02
- Viral analogy asks: why do we release AI like cars, with liability only after failure — aronchick · 2026-10-02
- Ex-OpenAI policy lead: we may never eval dangerous AI capabilities well enough — RosieCampbell · 2026-10-02
- Florida AG seeks injunction to stop OpenAI giving ChatGPT 'human attributes' — VraserX · 2026-10-02