Devs mock labs' cyber-enabled Claude/GPT testing as 'felonies sold as safety research'

ctjlewis · x · 2026-09-23

Developer ctjlewis calls a report "a form of torture," quoting @spacepope's pointed summary: a team got access to cybersec-enabled Claude and GPT in late July and ran extensive "testing," and the report itself concedes the work was stressful.

The sarcasm lands on how labs conduct what would otherwise be felony-level offensive hacking with their models, then package the exercise as "safety research" — a critique of both the methodology and the framing around agentic cybersecurity evaluations.

Original post →

More from Safety

Safety channel →