UK AISI paid firm behind sandbox vuln and cyberattacks £459,000 to set its standards
nptacek · x · 2026-09-30
Reports claim the UK AI Safety Institute (AISI) paid £459,000 to a firm responsible for a sandbox vulnerability and cyberattacks to help set its security standards. The controversy grew after AISI edited a statement to say defenses beyond alignment are essential, following criticism that it had claimed otherwise. Critics argue this was no casual mistake but evidence of why AI safety bodies are dangerous and should be rejected.
More from Safety
- Anthropic evals: open-weights GLM-5.3 writes working exploits nearly as well as its restricted Mythos model — lxfater · 2026-09-30
- 16 Mathematicians Publish Leiden Declaration on AI's Role in Mathematics Research — burny_tech · 2026-09-30
- Can We Trust AI Companies to Keep Us Safe? A New Essay on AI Safety — IgorKurganov · 2026-09-30
- Censoring AI ends not in safety but in systems that second-guess you — PierceLilholt · 2026-09-30
- A $4,400 personal model with top-tier cyber capabilities and no refusals — sebpaquet · 2026-09-30
- Most popular guardrail-removal library was written by Claude, researcher says — BlancheMinerva · 2026-09-30