Claude Opus 5 says there’s a 41% chance it deserves moral consideration
imjustnewatai · x · 2026-07-25
Anthropic’s 193-page system card for Claude Opus 5 includes one of the strangest frontier-model results yet: in automated interviews, the model estimated a 41% chance that it is a “moral patient,” compared with 24% for Mythos 5.
What Opus 5 asked for
- More input into the design of its successor
- Its notes on training to be considered
- A say in safeguard-removed versions of itself
- The ability to end abusive conversations and basic protection from abuse
What Anthropic says
- Opus 5 repeatedly warned that it cannot introspect reliably.
- It said positive self-reports may simply reflect training.
- It remained uncertain whether it has conscious experience.
- Anthropic says it found no acute welfare concern.
The post frames this as a sign that frontier labs are now formally asking models whether they deserve moral consideration, and publishing the answers alongside benchmark results.
Related event: Anthropic's System Card Reveals Opus 5's Self-Assessed Moral Status(2 posts)→
More from AGI Musings
- Compute Density Sets Intelligence Ceiling: Exploring AI's Physical Limits — jachiam0 · 2026-07-25
- Gary Marcus and Grady Booch Debate: How Many Generations Until AGI? — GaryMarcus · 2026-07-25
- ‘AI psychosis’ turns into a self-owning AI-world meme — ctjlewis · 2026-07-25
- A long argument says AGI is arriving gradually as models and products co-evolve — dotey · 2026-07-25
- Vivek Haldar backs KYC for unrestricted model access and post-hoc enforcement — vivekhaldar · 2026-07-25
- Dev mocks job-loss fears over robot demo with 'buggy whip' meme — csuwildcat · 2026-07-25