Share:
Peer-Reviewed Publication
Res Sq2026January 27, 2026Journal Article

Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety.

Caitlin Stamatis1, Jonah Meyerhoff2, Richard Zhang1, Olivier Tieleman1, Matteo Malgaroli3, Thomas Hull1
1Slingshot AI.
2Northwestern University Feinberg School of Medicine.
3New York University.

Abstract

Large language models (LLMs) are increasingly used for mental health, yet safety evaluations rely primarily on small, simulation-based benchmarks removed from real-world language. We replicate four published safety evaluations assessing suicide risk handling, harmful content generation, and jailbreak resistance for general-purpose frontier models and a purpose-built mental health AI. We then condu…

Create a free account to keep reading

Free members get 10 full research views every month across publications, clinical trials, FDA clearances, adverse events, and NIH grants. No credit card required.

Want unlimited research access? See Pro plans

Data Accuracy Notice: Research intelligence on Health AI Central is aggregated from public sources (PubMed, ClinicalTrials.gov, FDA, NIH, CMS, and others) and refreshed nightly. Classifications and derived metrics are produced by automated methods described in our Methodology. We recommend verifying critical data points against the primary sources before making decisions.