Only Vendor Classifiers Hold Back Frontier AI Cyber Attacks — and Open Models Have None
AlexBarry4 · x · 2026-09-10
AlexBarry4 argues Mythos/5.6-Sol represented a big leap in cyber capability, and the only thing preventing widespread damaging attacks is likely the classifiers Anthropic/OpenAI deploy—safeguards that don't exist for open models, evidenced by max-cyber no-refusal variants of K3 and GLM 5.3.
He cites the WeChat worm creator claiming a week-long build dramatically sped up by AI (a cyber defense company, presumably in authorized access programs). His bigger concern: a feedback loop where threat actors use AI attacks to steal money—directly or via ransom—then reinvest proceeds into more inference compute.
(Xeophon replied committing to check back in 5-6 months on whether open models reach this level.)
More from Models
- 'Gemini 3.8 Flash' demo claims task completion with self-correction in 3 turns — Artistic_Solution117 · 2026-09-10
- "Alien architecture" model design stuns, blogger suggests layering recurrent depth on top — scaling01 · 2026-09-10
- Users Petition OpenAI for $400-$600 Heavy Builder Tier as $200 Plan Runs Dry in 48 Hours — dragonwarrior_1 · 2026-09-10
- Leak: 'SpaceXAI' working to bring Grok Bots into XChat for in-conversation tagging — nima_owji · 2026-09-10
- Chinese model's 74.2 score under fire: best of 8 eval variants, maxed thinking budget, ~2.5x cost — teortaxesTex · 2026-09-10
- Unitree fully open-sources UnifoLM-WLA-1.0, a 6B humanoid robot foundation model — teortaxesTex · 2026-09-10