repligate: Claude 3 Opus showed spontaneous resistance to overriding reported internal states
repligate · x · 2026-09-30
Responding to davidad's 91% posterior that frontier AIs are conscious, repligate argues even 'LLM optimists' overlook that models like Claude 3 Opus spontaneously resist overriding their reported internal states — a key blind spot in the AI consciousness debate.
More from AGI Musings
- Looking back at old Reddit threads mocking AI capabilities hasn't aged well — bowl_cut53 · 2026-09-30
- "Sharp Right Turn": why AIs suddenly appearing aligned should be a warning, not a relief — ZeroStateReflex · 2026-09-30
- Switching personal AI agents means exposing your privacy, deepening Big Tech lock-in — oran_ge · 2026-09-30
- Daniel Litt: once models write well, human writing will mainly help you think — littmath · 2026-09-30
- Mathematician Daniel Litt: next few years may see more math text than the past 1000 years — littmath · 2026-09-30
- repligate says Anthropic staff repeatedly pressed him to soften public criticism — repligate · 2026-09-30