New "systems are thinking" guardrail message hints at OpenAI human review
Darpinian · x · 2026-10-05
A user spotted a novel guardrail rejection message—"systems are thinking"—from OpenAI and speculates the company may be using human reviewers to adjudicate uncertain safety cases when automated guardrails are unsure.
More from Models
- Dev claims mystery system hits 100% on agentic benchmarks, can't explain how — examachine · 2026-10-05
- Dev slams Opus for refusing to fix security holes that GLM5.3 found — lucasmeijer · 2026-10-05
- Whistle, an on-device speech recognition model, trends on Hugging Face — Cactus-Compute · 2026-10-05
- Aleph Alpha's Kolibri 78B is now free to try online — EveYogaTech · 2026-10-05
- Near-identical image scores, huge gaps: AI denoising must serve science, not looks — bravo_abad · 2026-10-05
- Decision Index: 70 open reproductions of Jev's Decision Model, one benchmark — bibryam · 2026-10-05