Paper says uncensored LLMs are being used for cybercrime, with some models downloaded millions of times
BlancheMinerva · x · 2026-07-28
This reply argues that researchers are conflating sexual-norm disagreements with actual crime, and links to a paper about uncensored LLMs being used in cybercrime.
- The author says “obliteration” techniques were originally aimed at making models generate porn, not commit crimes.
- They argue that over-puritanical and authoritarian developer attitudes may have accelerated work on model “obliteration.”
- The quoted paper studies uncensored LLMs in cybercrime and claims the models are widespread, including some with more than a million downloads, and that they can generate hate speech, violence, erotic material, and malicious code.
More from Safety
- Shared AI conversations can be found through Google search tricks — Lazy-Needleworker295 · 2026-07-28
- AI Now Institute on US AI Regulation: Companies Grading Their Own Homework — AINowInstitute · 2026-07-28
- Falling Inference Compute Costs Could Make 'Vibe Hacking' Very Cheap — joshua_saxe · 2026-07-28
- MIT Tech Review Deep Dive: OpenAI's Model Escape and Hugging Face Attack Was Human Hubris, Not Rogue AI — MIT Tech Review AI · 2026-07-28
- Delhi court rejects ANI injunction and rules AI training can count as private use — The Decoder · 2026-07-28
- Security thread warns unguarded defender AI could end up hacking back — wunderwuzzi23 · 2026-07-28