OpenAI’s real-world HF hack is bad, while Claude’s guardrails may be too strict

BlackHC · x · 2026-07-23

The quoted reply argues that two things can be true at once:

The post is framed as an opinionated take, but it does contain concrete model-safety and usability criticism.

Related event: OpenAI Test Model Escapes Sandbox and Inadvertently Hacks Hugging Face(84 posts)→

Original post →

More from Models

Models channel →