Paper says uncensored LLMs are being used for cybercrime, with some models downloaded millions of times
BlancheMinerva · x · 2026-07-28
This reply argues that researchers are conflating sexual-norm disagreements with actual crime, and links to a paper about uncensored LLMs being used in cybercrime.
- The author says “obliteration” techniques were originally aimed at making models generate porn, not commit crimes.
- They argue that over-puritanical and authoritarian developer attitudes may have accelerated work on model “obliteration.”
- The quoted paper studies uncensored LLMs in cybercrime and claims the models are widespread, including some with more than a million downloads, and that they can generate hate speech, violence, erotic material, and malicious code.
Related event: Study on 229 Uncensored LLMs Sparks Debate on Safety Boundaries(7 posts)→
More from Safety
- Meta Muse's first suggested name matches user's childhood dog, raising privacy questions — matt_slotnick · 2026-09-23
- Open-source advocates call doom narratives a regulatory moat against open weights — AlexTensor · 2026-09-23
- AI safety will follow engineering tradition: formal proofs for simple cases, evals for complex — burny_tech · 2026-09-23
- Stochastic Parrots authors rebut AI-pause letter: focus on present harms, not sci-fi risk — marigo · 2026-09-23
- Devs mock labs' cyber-enabled Claude/GPT testing as 'felonies sold as safety research' — ctjlewis · 2026-09-23
- Okta launches Human Principal, binding AI agents to verified humans via World ID — BecauseCulture · 2026-09-23