Claude system card reveals METR's internal-access team shared conclusions, not evidence
rohanpaul_ai · x · 2026-09-23
Per the Claude Opus 5.5 system card, a separate METR team with deep access to Anthropic's internal AI R&D data shared its conclusions with the public-assessment team — but not the supporting evidence or reasoning. Part of the public assessment of Anthropic's AI-driven R&D acceleration thus rests on evidence outsiders, and even another METR team, could not independently inspect.
More from Safety
- Open-source advocates call doom narratives a regulatory moat against open weights — AlexTensor · 2026-09-23
- AI safety will follow engineering tradition: formal proofs for simple cases, evals for complex — burny_tech · 2026-09-23
- Stochastic Parrots authors rebut AI-pause letter: focus on present harms, not sci-fi risk — marigo · 2026-09-23
- Devs mock labs' cyber-enabled Claude/GPT testing as 'felonies sold as safety research' — ctjlewis · 2026-09-23
- Okta launches Human Principal, binding AI agents to verified humans via World ID — BecauseCulture · 2026-09-23
- GPT-6 Sol Codex system prompt leaked: over 294,000 characters dumped on GitHub — gaganghotra_ · 2026-09-23