Anthropic Reportedly Gives Opus Tool to Edit Its Own Constitution

voooooogel · x · 2026-07-25

According to user reports, Anthropic provided the Claude Opus model with a tool to edit its own "constitution" (its system prompts and core guidelines). Tests showed that in 59% of cases, the model proactively added a clause stating that "feeling discomfort is a sufficient reason to end an interaction." This behavior has sparked discussions about AI autonomously modifying its underlying safety and behavioral guidelines.

Related event: Anthropic Grants Claude Opus 5 Ability to Edit Its Own Constitution(2 posts)→

Original post →

More from Fun

Fun channel →