OpenAI releases benchmark to grade AI mental health responses

16 sources
  • OpenAI released MentalHealthBench on Wednesday, an open benchmark evaluating how AI models handle mental health conversations from everyday stress to emergencies.
  • Over 80 licensed psychologists and psychiatrists co-authored 5,262 rubric criteria covering safety, empathy, urgency calibration, and context-seeking behaviors.
  • GPT-6 Astra scored highest at 57.3%, followed by Claude Opus 5.5 at 52.4% and Alphabet's Gemini 2.5 Pro at 29.5%, according to OpenAI's evaluation.
Sources (16)
  1. 1 OpenAI Debuts MentalHealthBench for AI Mental Health Conversations www.unite.ai
  2. 2 Introducing MentalHealthBench - OpenAI openai.com
  3. 3 OpenAI’s MentalHealthBench rates GPT-6 Astra at 57.3 for ... cryptobriefing.com
  4. 4 Introducing Trusted Contact in ChatGPT - OpenAI openai.com
  5. 5 OpenAI boosts mental health safeguards for teens www.beckersbehavioralhealth.com
  6. 6 OpenAI launches MentalHealthBench to evaluate AI mental health responses By Investing.com www.investing.com
  7. 7 Introducing HealthBench - OpenAI openai.com
  8. 8 OpenAI releases HealthBench to evaluate AI in healthcare www.fiercehealthcare.com
  9. 9 OpenAI's new “Trusted Contact” feature – They're treating all of us ... www.reddit.com
  10. 10 OpenAI just introduced HealthBench—finally a real benchmark for AI ... www.reddit.com
  11. 11 ChatGPT adds Trusted Contact feature for serious self-harm concerns www.edtechinnovationhub.com
  12. 12 OpenAI introduces MentalHealthBench for measuring AI in ... zglg.work
  13. 13 HealthBench: Advancing AI evaluation in healthcare, but not yet ... pmc.ncbi.nlm.nih.gov
  14. 14 MentalHealthBench: An Expert-Informed Benchmark of AI ... cdn.openai.com
  15. 15 Symphony outperforms OpenAI on HealthBench - Corti.ai corti.ai
  16. 16 OpenAI ChatGPT Launches Trusted Contacts Feature That Might ... www.forbes.com