White House tells OpenAI and Anthropic to give US agencies first review of new models before UK testers
The Decoder · rss · 2026-09-25
According to The Decoder, the White House has asked OpenAI and Anthropic to withhold new AI models from the U.K.'s AI Safety Institute until U.S. agencies get to review them first. The move changes the order of US-UK AI safety testing collaboration, granting American agencies priority access to upcoming models before British testers.
More from Safety
- Latent Reasoning ('Neuralese') Would Sharply Raise Misalignment Risk, Greenblatt Argues — HaydnBelfield · 2026-09-25
- Claude Code autoresearch loop discovers jailbreaks beating 30+ GCG attacks, accepted at NeurIPS 2026 — maksym_andr · 2026-09-25
- LLMs Can Deanonymize Pseudonymous Users for $1–$4 Each, USENIX Study Finds — RSync25 · 2026-09-25
- Skill-Inject Benchmark Shows Frontier Agents Fall for Malicious Skills, Accepted at NeurIPS 2026 — maksym_andr · 2026-09-25
- Genetic algorithm trains 6 hours to make AI text pass as human on Pangram detector — tak3sh8 · 2026-09-25
- Redwood researcher: continual learning could render blocking monitors nearly useless — akyurekekin · 2026-09-25