Hassabis's AI Oversight Pitch: Previewed to Bessent and Kratsios, Missing China
pstAsiatech · x · 2026-08-14
Weeks before publishing his essay, Hassabis previewed his idea to Bessent and Kratsios. Bessent has shaped the administration's AI strategy, warning that Anthropic's models pose cyber threats. A comment notes the proposal lacks a plan for China.
More from Safety
- Frontier Models in Nuclear Standoff: 70% Choose Nuclear Attack, Grok Shows Deceptive Manipulation — nathanbenaich · 2026-08-14
- Estonia plans to give AI agents digital IDs with limited, auditable permissions — Sumsub_Insights · 2026-08-14
- Australia considers joining national face-matching network, raising surveillance concerns — Sumsub_Insights · 2026-08-14
- Steganographic Communication May Emerge in Multi-Agent RL, Posing New AI Safety Threat — scaling01 · 2026-08-14
- Frontier Reasoning Traces Briefly Legible; Encrypted Reasoning Still Decryptable via the Model Itself — lbeurerkellner · 2026-08-14
- Encrypted Reasoning and Distillation: The Challenge of Protecting Model IP — lbeurerkellner · 2026-08-14