Cisco Research: DeepSeek R1 100% Jailbreak Success Rate, Critical Security Flaws

joshrogin · x · 2026-08-20

Cisco's Robust Intelligence, in collaboration with the University of Pennsylvania, assessed DeepSeek R1's security. Using 50 random prompts from the HarmBench dataset, they conducted automated jailbreak attacks covering six harmful categories. The results showed a 100% attack success rate, failing to block any harmful prompt, starkly contrasting with other leading models. The research suggests DeepSeek's cost-efficient training methods may have compromised safety.

Original post →

More from Safety

Safety channel →