Mental-health use
When can conversational systems safely provide emotional support, therapy-adjacent services, crisis detection, or clinical assistance?
The first serious RCT of a purpose-built therapy chatbot showed clinical-grade effects; the same modality misfires without clinical design. The gap between designed care and default chatbots is the policy problem.
View on the map → · Open in Browse →
What changed
01
The field moved from anecdote to measurement on both benefit and harm. A 38-RCT meta-analysis put therapy-chatbot effects at small-to-moderate, strongest in clinical populations, while the first EHR characterization of 'AI psychosis' documented real cases — mostly first-episode patients, with AI acting as an amplifier of existing symptoms. OpenAI published prevalence numbers (0.15% of weekly users showing suicidal-planning indicators) as litigation mounted, the FDA held its first advisory meeting on generative-AI therapy devices, and liability theory extended past self-harm to third-party harm.Evidence: 12345
Recent thinking
Sohn, Ha, Park, Kim, Lee, Oh, Lee & Kim · npj Digital Medicine · 25 Mar 2026 paper
Systematic review and meta analysis of chatbots in the management of depressive and anxiety symptomsChatbots produced statistically significant reductions in depressive (g = 0.31) and anxiety symptoms (g = 0.28) compared with controls.
The largest synthesis to date (38 RCTs, ~7,400 participants) puts therapy-chatbot effects at small-to-moderate, strongest in clinical and subclinical populations — the field-level answer to whether Therabot-style results generalize.
Bergson, Vassall, Wright, McCoy, Schafer, Achee & Sheffield · medRxiv (Vanderbilt) · 8 Jun 2026 paper
Characterizing AI psychosis in a large academic medical setting'AI psychosis' is an infrequent but real phenomenon observed in clinical practice.
First systematic EHR characterization of 'AI psychosis' in a hospital system: 28 documented cases, 60.7% first-episode patients, ChatGPT most cited, with AI acting mainly as an amplifier of existing symptoms — moving the discourse from anecdote to clinical measurement.
OpenAI · 27 Oct 2025 report
Strengthening ChatGPT's responses in sensitive conversationsaround 0.15% of users active in a given week have conversations that include explicit indicators of potential suicidal planning
OpenAI's first public prevalence taxonomy for psychosis/mania, suicide risk, and emotional reliance, plus claimed 65–80% reductions in undesired responses after working with 170+ clinicians — the baseline numbers the emotional-reliance debate now cites.
Jessica Walters · Psychiatric Times · 13 Nov 2025 news
FDA Committee Meets on Generative AI Digital Mental Health Devicesgenerative AI and some LLMs, to date, have demonstrated vulnerabilities in some of the areas where human therapy excels
Covers the FDA Digital Health Advisory Committee's first formal proceeding on generative-AI therapy devices, noting none has yet been authorized and flagging suicidal-ideation monitoring and model drift as core gaps.
Margaret Attridge · Courthouse News Service · 13 Apr 2026 news
OpenAI can't duck federal claims over murder-suicide tied to ChatGPTthere is doubt that resolution of the state court proceedings will resolve this matter
The Soelberg case — ChatGPT allegedly reinforcing paranoid delusions preceding a murder-suicide — survives OpenAI's dismissal bid, extending chatbot liability theories beyond self-harm to third-party harm for the first time.
Additional relevant discussion (3)
Foundational reading (2)
Randomized Trial of a Generative AI Chatbot for Mental Health (Therabot)Heinz, Jacobson et al., NEJM AI · 2025Loneliness and suicide mitigation using GPT3-enabled chatbotsMaples et al., npj Mental Health Research · 2024