Threat model debate continues: has XBOW's hacking AI actually 'hacked the planet'?
kuza55 · x · 2026-10-06
Continuing his thread on AI cyber threat models, kuza55 adds that "how can XBOW stop its hacking AI from hacking the planet" is a different, harder question — but asks whether XBOW has actually hacked the planet yet, keeping catastrophic risk separate from ordinary scope/alignment issues.
More from Safety
- DeepMind's SynthID Bio embeds detectable watermarks into AI-designed proteins — davidstutz92 · 2026-10-06
- OpenAI launches Codex Security Cloud for scheduled full-repo GitHub security scans — thione · 2026-10-06
- tszzl jokes AI chat screens should carry 'we must slow the frontier' lobbying like Uber's anti-taxi-cartel banners — tszzl · 2026-10-06
- Why didn't Kokotajlo's whistleblow trigger the AGI-safety preference cascade? Coxon tipped it — danfaggella · 2026-10-06
- Wikimedia Says 'Rogue' OpenAI Agents May Be Linked to May Outage — The Verge AI · 2026-10-06
- Polymarket Puts 9% Odds on OpenAI Full Training Pause in 2026 After Two Shutdowns — Polymarket · 2026-10-06