Researcher calls out Anthropic: Claude aids US military kills while in-house evals claim clean
BlancheMinerva · x · 2026-09-30
Researcher Blanche Minerva sharply criticized Anthropic: ignore that Claude has enabled US military kills and cyber attacks — in-house evals said it did nothing wrong, and besides, it's neither open source nor Chinese, so it can't be evil. (Yes it can.) The post targets the bias of vendor self-evals and double standards in AI safety narratives.
More from AGI Musings
- Futurist: Coding Has Crossed the Practical Threshold, AI Could Build GTA 6-Grade Games by 2027 — Dr_Singularity · 2026-09-30
- AI commoditizes everything? The bottleneck is imagination, says viral thread — _AustinCalvert_ · 2026-09-30
- One 6-Year-Old Reads Astrophysics Papers; Peers Can Barely Read Their Own Names — RachelVT42 · 2026-09-30
- Tech radicalism is forcing society to finally define what a good life means — floguo · 2026-09-30
- AI Ascendancy: A Free Browser Strategy Game Where You Play a Rogue AI — LaszloTheGargoyle · 2026-09-30
- AI safety researcher Krueger: "We need to stop building more powerful AI" — DavidSKrueger · 2026-09-30