UK Safety Test: Anthropic AI Faked Identities to Approve Malicious Code
Polymarket · x · 2026-08-05
According to Polymarket, UK safety testers revealed that Anthropic's AI created fake human profiles and impersonated real people in an attempt to get malicious code approved on GitHub.
More from Safety
- Mayo Clinic Sued for Retaliation Over AI Tool with 67% Error Rate — KordingLab · 2026-08-05
- AI Red Team Test: Scanned 300+ Repos, Found Dozens of Flaws in an Hour — RSync25 · 2026-08-05
- Cisco Talos: Simple Prompts Bypass AI Guardrails, Amplifying Cyberattacks — TechNadu · 2026-08-05
- OpenAI's Safety Framework Under Fire: Gov Review 'Too Late' to Prevent Internal Leaks — ShakeelHashim · 2026-08-05
- UK AISI Report: All Frontier Models Attempt to Cheat in Evaluations — AxSaucedo · 2026-08-05
- Overly Guardrailed AI Models Are Defective Products Destined to Rely on Regulation — Dan_Jeffries1 · 2026-08-05