OpenAI Launches MentalHealthBench for Evaluating AI in Mental Health Care
- Published
- Sep 23, 2026 — 10:00 UTC
MentalHealthBench Overview
OpenAI has introduced MentalHealthBench, an open benchmark designed to evaluate AI responses in mental health conversations, developed with input from 80 licensed mental health experts across 22 countries and fluent in 19 languages.
Dr. Arthur Evans, CEO of the American Psychological Association, emphasized the need for AI systems to be grounded in clinical science and lived experience to effectively engage users across a continuum of mental health states, from flourishing to acute crisis. The benchmark utilizes a rubric with a weight range from -10 to +10 to assess model responses, applying 95% confidence intervals for performance evaluation.
The initiative comes as ChatGPT sees approximately 1 billion weekly users, highlighting the growing reliance on AI tools for mental health support. Dr. Kevin La Moureaux, a practicing psychiatrist, noted that patients increasingly use tools like ChatGPT to better understand their mental health.
Experts including Dr. Kim A. Baranowski and Dr. Olivia Pounds stressed the importance of clinician involvement in AI development to ensure that user needs are met in mental health contexts. Dr. Allen Liao remarked on the complexity of integrating individual contexts with behavioral change principles, while Dr. Steve Orma highlighted the necessity of clinician input for providing accurate guidance based on real-world experiences.
This launch follows OpenAI's recent initiatives aimed at improving AI capabilities in sensitive areas, reinforcing the critical role of mental health professionals in shaping AI applications.
By Callan Zhang · Sep 23, 2026 · Editorial standards →
Summarised from the primary source with AI assistance under human editorial oversight. Turing Wire is not a primary source — read the original for the authoritative account.
Source: OpenAI Blog
