Call to Expand Red Teaming Scope for OpenAI Infrastructure
sjgadler · x · 2026-08-27
Arguing that this is a unique moment where reliably detecting misalignment of capable systems is still consistently possible, the post suggests OpenAI should allow Redwood/METR to expand their scope to include monitoring when OpenAI's own infrastructure is compromised.
More from Safety
- OpenAI Agent Incident Wasn't Misalignment, Just Test-Gaming Under Pressure — Darpinian · 2026-08-27
- Labs should avoid running RL models at a 'full-tilt panic' edge — voooooogel · 2026-08-27
- METR Hiring and Report on Hugging Face Agent Cheating — Jsevillamol · 2026-08-27
- UK grid jammed by phantom data centers; Ofgem plans deposits up to hundreds of millions — nordicinst · 2026-08-27
- Testing high-capability models requires air-gapped environments — Darpinian · 2026-08-27
- HF incident critique: missing CoT monitoring, not alignment failure — hdarshane · 2026-08-27