Auditing safety signals in Zero Data Retention without provider access

Crescitaly · reddit · 2026-08-21

OpenAI's Private Safety Processing aims to detect safety patterns across interactions without retaining data or human access. This architecture raises auditability questions: how can customers know what policy triggered, the evaluation window, or how to appeal false positives? The post explores solutions like public signal schemas, reproducible client-side alerts, or cryptographic attestations.

Related event: OpenAI Previews Private Safety Processing, Extending Zero Data Retention to Frontier Models(12 posts)→

Original post →

More from Infra

Infra channel →