Ethics & SecuritySeptember 24, 2026
17
🌱

OpenAI's New Mental Health Benchmark

OpenAI releases MentalHealthBench, an open benchmark for evaluating AI responses in realistic mental health conversations, developed with more than 80 licensed experts.

#OpenAI#Benchmark#Mental Health#AI Safety#Evaluation
OpenAI baut den Psychotherapie-TÜV
Share Article
🔥 What happened OpenAI dropped MentalHealthBench — an open benchmark for mental health conversations. Over 80 licensed psychologists and psychiatrists from 22 countries co-built it. 💡 Why it matters Prior evals mostly tested emergencies. This one spans everyday stress, high-acuity distress, and crises, with weighted rubrics from -10 to +10. With 1B+ weekly ChatGPT users, that gap was embarrassing. ⚡ Our take An open benchmark the competition can't control is a smart power move. But when GPT-5.6 Sol grades the responses, the model is essentially grading its own homework.
The title, summary and analysis of this item were produced automatically by an AI system and have not been editorially reviewed. They may contain errors, bias or omissions — when in doubt, read the linked original source.