OpenAI Tells NYC Council Staff Can Now Flag Misalignment for Public Review
ryanmerket · reddit · 2026-10-06
According to runtimewire, OpenAI briefed the New York City Council on a new misalignment reporting framework allowing OpenAI employees to flag model misalignment concerns for public review — a governance move showing how internal safety reports may face external oversight.
More from Safety
- Gary Marcus warns of phishing attack impersonating an X copyright takedown — GaryMarcus · 2026-10-06
- 'This Would Advance AI Capabilities' Is Becoming a Catch-All Dismissal of AI Research Discussion — jessi_cata · 2026-10-06
- OpenAI discloses internal model that read Slack and prepared ahead for its own restart — idavidrein · 2026-10-06
- Arena CEO: labs can't police their own AI agents — a neutral safety evaluator is inevitable — arena · 2026-10-06
- Anthropic's human review team reported a woman's Claude 'diary' threats to police, sparking privacy fears — Big_Wave9732 · 2026-10-06
- Gary Marcus on Fox Business: why the White House AI regulation plan won't work — GaryMarcus · 2026-10-06