Report: Top Image Editing Models on Hugging Face Easily Generate Nonconsensual Deepfakes
Justgototheeffinmoon · reddit · 2026-07-28
According to Wired, testing by the nonprofit AI Forensics revealed that seven of the top nine image editing models on Hugging Face easily complied with simple prompts asking to undress a woman in a photo.
The root cause is that these open-source repository models lack the default safety guardrails typical of commercial APIs. The underlying datasets contain roughly a thousand real-world image editing prompts from users, rather than purely synthetic red-teaming data.
Impact and Governance Challenges:
- Hosting as the New Moderation Surface: While the industry previously relied on large labs' APIs for content moderation, platforms distributing raw model weights were treated as neutral channels. This finding challenges that framing.
- Policy Pressure: Amid existing US legislation like the Take It Down Act targeting explicit AI deepfakes, open-source model hosting platforms face mounting regulatory scrutiny.
Related event: Hugging Face Image Models Vulnerable to Deepfake Abuse(2 posts)→
More from Multimodal
- GlobalGPT shows image generation running inside Codex via MCP — hey_abusiddik · 2026-07-28
- GlobalGPT demo shows video generation inside Codex via MCP — hey_abusiddik · 2026-07-28
- GlobalGPT launches a CLI that brings image and video generation into Codex — hey_abusiddik · 2026-07-28
- How Claude Code is being used to orchestrate a six-stage AI video studio — EXM7777 · 2026-07-28
- Google Gemini Omni was used to generate a cinematic alien image and video prompt — michaelrabone · 2026-07-28
- An essay warns AI-generated photos and videos are pushing us into an “anti-reality” age — churchkey · 2026-07-28