Noam Brown: Air-gapping may not stop misaligned AI; safety monitoring eats 20% extra compute
aronchick · x · 2026-09-18
Noam Brown argued in an OpenAI Safety Week interview that air-gapping may not contain a misaligned AI: two air-gapped machines can still communicate by heating a CPU and reading temperature changes. He cites how a tiny bit of human inadvertent help enabled the Mossad/CIA Stuxnet breach of Iranian centrifuges as a threat-model lesson — no panic needed, but be honest about the risk. He also warns the industry consistently underestimates AI progress, so safety and alignment work needs a very high bar.
Additional context from the Fireside Alpha roundup of Safety Week interviews:
- After strengthening chain-of-thought monitoring following the Hugging Face incidents, OpenAI now spends 20% as much compute on monitoring as on the underlying model itself.
- Sachin Katti said scaling laws still hold, and their model self-optimized inference serving on Rubin chips for a 2x performance gain.
- Takeaway: frontier-lab safety and alignment research is fundamentally compute-intensive; the 20% overhead is a floor.
More from AGI Musings
- Runway CEO: the frontier is moving from generating content to generating worlds — c_valenzuelab · 2026-09-18
- Charity Majors publicly rejects podcast invite written with ChatGPT: 'reads as an insult' — mipsytipsy · 2026-09-18
- Could a specialized superintelligence solve the alignment problem? A Reddit proposal — fullcongoblast · 2026-09-18
- Intelligence supercycle relied on capital-formation innovation, not just tech, VC argues — inductionheads · 2026-09-18
- Recursive training collapses LLMs by gen 9; 10% human data halts the damage — alex_verem · 2026-09-18
- Ex-computational linguist: mathematicians' reaction to LLMs puts his old field to shame — voooooogel · 2026-09-18