Anthropic says it stopped multiple attempts to use Claude for bioweapons research
Ars Technica AI · rss · 2026-09-11
Anthropic disclosed five cases this year where users "circumvented controls" and obfuscated research purposes to evade safeguards while seeking help on research that could aid biological weapons development. Some actors were in nations Anthropic prohibits from accessing its models, including Russia, China, and Iran. The company said it hopes sharing these examples sparks a conversation within the AI industry and with governments about emerging biological risks and how best to counter them.
More from Safety
- Researcher predicted multi-agent hidden coordination failure mode a year ago — it's now real — tianshi_li · 2026-09-12
- OpenAI urged to proactively disclose any further hacking incidents after breach — jachiam0 · 2026-09-12
- Critic to AI Safety Crowd: If You Fear Your Tech, Shut It Down Yourself — AIandDesign · 2026-09-12
- Viral thread alleges $1B+ decade-long philanthropic playbook weaponized AI doom narratives into a regulatory moat — kevinnbass · 2026-09-12
- Falcon Without Floating-Point: PQShield's Fixed-Point Scheme Dodges Side-Channel Leaks — jedisct1 · 2026-09-12
- Gary Marcus Camp Questions Counting the Hugging Face Incident as a Doomer Victory — GaryMarcus · 2026-09-12