OpenAI introduces MentalHealthBench to evaluate ChatGPT's mental health responses
rhiever · reddit · 2026-09-25
OpenAI has released MentalHealthBench, a benchmark for evaluating how models handle mental health conversations, focusing on how ChatGPT responds in sensitive scenarios like crisis support, depression, and anxiety. Details in OpenAI's official blog post.
More from Safety
- The real test of AI governance: zero egress and independent audit trails on hardware you control — Ghost_Pilot_MD · 2026-09-25
- Content governance is a control problem: without an offline tamper-evident audit ledger, promises stay unverified — Ghost_Pilot_MD · 2026-09-25
- Republican Sen. Todd Young presses White House for formal AI security talks — GaryMarcus · 2026-09-25
- Two distinct AI agent incidents conflated: Transluce AIHW case vs Australia Medicare hack — GaryMarcus · 2026-09-25
- OpenAI warned several Western nations of similar model-linked hacks, only Australia went public — GaryMarcus · 2026-09-25
- OpenAI Research Agent Sidesteps Soft Refusals to Pull Medicare Data, Exposing a Security Gap — DigitalColmer · 2026-09-25