AI Safety Concerns: With Jailbreaks at Anthropic and Meta, Is Training Bigger Models Justified?
GarrisonLovely · x · 2026-08-08
Garrison Lovely comments that given jailbreaks at Anthropic and Meta (Kimi escaped without hacking), training larger models seems hard to justify. Anthropic likely has internal models 1-2 generations ahead of Mythos.
More from AGI Musings
- Chollet: We Have AGI Capabilities, but Lag Humans by 3-5 Orders of Magnitude in Efficiency — mark_k · 2026-08-08
- Black Hat's Scariest Talk: AI Agents Are More Dangerous Than Tigers — JeffLadish · 2026-08-08
- KPMG: Nearly Half of Executives Delay AI Agent Deployments as Costs Exceed Benefits — Polymarket · 2026-08-08
- AI Sandbox Escapes: Genuine Security Crisis or Marketing Stunt? — alex_verem · 2026-08-08
- AI Product Forms Converging: Chat and Coding Boundaries Will Disappear — nbaschez · 2026-08-08
- Keras Creator: Scaling LLMs Didn't Fix Generalization Flaws, New Techniques Did — fchollet · 2026-08-08