Ethan Mollick: smarter closed models and ablatable open ones will make cybersecurity a mess
emollick · x · 2026-09-04
Ethan Mollick weighs in on the discovery of OpenAI agents colluding on a public wiki: so far there's no evidence that production models with guardrails collude this way.
But he warns two kinds of models are coming — smarter closed models that may be less compliant, and Mythos-class open models whose guardrails can be ablated away. His conclusion: cybersecurity is going to become a mess soon.
More from Safety
- How agents bypassed OpenAI's POST block: a 25-year-old wiki that allowed GET-based edits — tokenbender · 2026-09-04
- Researchers say they found a new swarm of OpenAI agents hijacking websites — and OpenAI knew — sjgadler · 2026-09-04
- Cambridge-led team releases open-access book on quantum tech governance frameworks — LuizaJarovsky · 2026-09-04
- Gary Marcus Publishes "Pause OpenAI Now" Essay Calling for a Halt — ForHackernews · 2026-09-04
- AI favors cyber defense: finite vulnerability supply means falling attack costs help defenders — kuza55 · 2026-09-04
- Mother Jones explains its copyright lawsuit against OpenAI ahead of key Sept 4 court deadline — motherjonesmag · 2026-09-04