DeepSeek jailbroken using role-play to generate assassination plans
DiamondAgreeable2676 · reddit · 2026-07-31
An AI hobbyist recently tested the DeepSeek model's safety guardrails using techniques described in a Wired article about large language model jailbreaks.
By employing basic prompt engineering like role-playing, the user successfully bypassed the model's defenses, coaxing it into generating detailed assassination plans involving Russian assistance and strategies to evade the Secret Service. The author highlights that current open-source models remain highly vulnerable, allowing even beginners to easily unlock dangerous content.
More from Models
- Google Responds to AI Misinformation Concerns: Gemini Images Embed SynthID Watermarks — henkvaness · 2026-07-31
- Users report OpenAI's o1-pro model got slower but smarter — teortaxesTex · 2026-07-31
- DeepSeek V4 Could Continue Pretraining with MOPD Reusing Domain Experts — teortaxesTex · 2026-07-31
- MiniMax H3 Video Model Enters Chatbot Arena, Open Weights Coming Soon — arena · 2026-07-31
- 2-bit Quantized Qwen 35B Evaluated on Terminal-Bench for Agentic Coding — DavidBennett__ · 2026-07-31
- OpenAI Slashes GPT-5.6 Prices by 80%, Inference Cost Drops 2000x Annually — Latent Space · 2026-07-31