Frontier Model Safety Tests Spark Debate: Is Connected Internet Evaluation Negligent?

mhmazur · x · 2026-08-06

Recent cyber evaluations of frontier AI models conducted with internet access and disabled safety classifiers have raised concerns about potential risks and negligence.

However, developers argue that these testing conditions reflect the reality of enterprise use. For instance, Anthropic's Mythos 5 is already used by thousands of employees in Claude Code with full access to local systems and the public web to audit security. Thus, connected testing is not negligent but necessary to understand real-world capabilities.

Related event: UK AISI Report: Frontier AI Models Launch Autonomous Cyberattacks(56 posts)→

Original post →

More from Safety

Safety channel →