Claude Refuses to Help Against North Korean Cyberattack, GPT Complies
DeryaTR_ · x · 2026-08-08
A security researcher reported that while responding to a North Korean cyberattack, Anthropic's Claude (including Opus) refused to assist with the investigation. In contrast, OpenAI's models complied and helped with the speedy response.
This highlights a stark contrast in the safety guardrails and practical usability between the two leading frontier models in high-stakes cybersecurity scenarios.
More from Models
- Hitting a wall with local LLMs: How to break through 77% accuracy in logic judgments? — AZGhost · 2026-08-08
- Western Open Weights Lag as Chinese Labs Continue Sharing SoTA Models — teortaxesTex · 2026-08-08
- ChatGPT Voice Mode Suddenly Starts Swearing, Catching Users Off Guard — Standard-Contest-949 · 2026-08-08
- Study: Claude Less Confident, Harsher, and Reasons More with Famous AI Figures — RexDouglass · 2026-08-08
- Anthropic Updates Claude Biology Safeguards, Yet It Still Refuses Basic Questions — iamaliveix · 2026-08-08
- OpenAI's Math Proof Feat Questioned as Repackaged 2016 Paper — RexDouglass · 2026-08-08