Anthropic: Opus 5.5 scores strongest alignment results to date
claudeai · x · 2026-09-23
Anthropic says Opus 5.5 is its first model since calling for frontier pacing, was externally evaluated by METR and Frontier Design before release, and achieved the strongest score to date on its automated behavioral audit alignment test.
More from Models
- Andriy Burkov: Codex Is Infinitely Faster Than Any Open-Weight Agent, But Priced Out of Reach — burkov · 2026-09-23
- Viral demo claims 'GPT-6' can drive browser Paint to draw, unverified — alexcovo_eth · 2026-09-23
- Opus 5.5-generated three.js spell demo wows with procedural VFX and sound — majidmanzarpour · 2026-09-23
- Early hands-on with rumored Opus 5.5 in Scenario's Blender plugin stuns users — repligate · 2026-09-23
- GPT-6 Sol's price cut means it should be compared to Sonnet, not Opus — TraditionalHome8852 · 2026-09-23
- User burns two OpenAI 20x accounts down to 13%, sizing up GPT-6 efficiency — jdjohnson · 2026-09-23