REVIEW 13 cited by
Benefits and Harms of Large Language Models in Digital Mental Health
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The past decade has been transformative for mental health research and practice. The ability to harness large repositories of data, whether from electronic health records (EHR), mobile devices, or social media, has revealed a potential for valuable insights into patient experiences, promising early, proactive interventions, as well as personalized treatment plans. Recent developments in generative artificial intelligence, particularly large language models (LLMs), show promise in leading digital mental health to uncharted territory. Patients are arriving at doctors' appointments with information sourced from chatbots, state-of-the-art LLMs are being incorporated in medical software and EHR systems, and chatbots from an ever-increasing number of startups promise to serve as AI companions, friends, and partners. This article presents contemporary perspectives on the opportunities and risks posed by LLMs in the design, development, and implementation of digital mental health tools. We adopt an ecological framework and draw on the affordances offered by LLMs to discuss four application areas -- care-seeking behaviors from individuals in need of care, community care provision, institutional and medical care provision, and larger care ecologies at the societal level. We engage in a thoughtful consideration of whether and how LLM-based technologies could or should be employed for enhancing mental health. The benefits and harms our article surfaces could serve to help shape future research, advocacy, and regulatory efforts focused on creating more responsible, user-friendly, equitable, and secure LLM-based tools for mental health treatment and intervention.
Forward citations
Cited by 13 Pith papers
-
The Impact of Security and Privacy Controls on Users' Emotional Engagement with Generative AI Chatbots
In a vignette study of 354 U.S. participants, deletion-based privacy controls outperformed all other controls in increasing willingness to engage with GenAI chatbots for emotional support, while technically complex co...
-
DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots
DelusionEval finds that AI chatbots show delusion-linked behaviors on real user transcripts and that longer conversation context increases the rate of some harmful responses.
-
Exploring User Security and Privacy Attitudes and Concerns Toward the Use of General-Purpose LLM Chatbots for Mental Health
Many users of general-purpose LLM chatbots for mental health misunderstand how their data is protected, conflating human-like empathy with accountability and undervaluing emotional disclosures as a privacy risk.
-
`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts
Academic-framing prompts bypass safety filters in most tested LLMs, turning prior self-harm and suicide intent into detailed actionable instructions.
-
Understanding Attitudes and Trust of Generative AI Chatbots for Social Anxiety Support
People with severe social anxiety symptoms report greater trust in and willingness to use GenAI chatbots, valuing emotional connection, while milder-symptom users emphasize technical reliability.
-
Evaluating an LLM-Powered Chatbot for Cognitive Restructuring: Insights from Mental Health Professionals
A GPT-4 chatbot followed cognitive restructuring steps for 19 users, but mental health experts flagged toxic positivity, advice-giving, and context misunderstandings.
-
Engagement and Disclosures in LLM-Powered Cognitive Behavioral Therapy Exercises: A Factorial Design Comparing the Influence of a Robot vs. Chatbot Over Time
A two-week factorial study with 26 students found that engagement and evaluative intimacy increased with a physical robot but decreased with a chatbot.
-
Human vs. LLM-Based Thematic Analysis for Digital Mental Health Research: Proof-of-Concept Comparative Study
GPT-4o with RISEN prompts can perform thematic analysis faster and cheaper, but humans still excel at child-code development, excerpt coding, and theme synthesis.
-
AI Chatbots for Mental Health: Values and Harms from Lived Experiences of Depression
People with lived depression experience who tried a GPT-4o chatbot prioritized five values: informational support, emotional support, personalization, privacy, and crisis management.
-
Expressing stigma and inappropriate responses prevents LLMs from safely replacing mental health providers
Current large language models show stigma and give clinically inappropriate responses to common mental health symptoms, so they should not be deployed as replacement therapists.
-
Interpersonal Theory of Suicide as a Lens to Examine Suicidal Ideation in Online Spaces
Using the Interpersonal Theory of Suicide as a lens, the authors classify 59,607 Reddit suicide-related posts into risk categories and find AI support responses are more coherent but less empathetic than human ones.
-
AI-Augmented LLMs Achieve Therapist-Level Responses in Motivational Interviewing
A custom prompt built from machine-learning-identified therapy behavior features improved GPT-4's motivational interviewing quality scores, though the model remained slightly below human therapists on the paper's own metric.
-
Harnessing Large Language Models for Mental Health: Opportunities, Challenges, and Ethical Considerations
A narrative review of how large language models might help and harm mental health care, concluding that ethical safeguards and human oversight are needed.
Discussion (0). Continue with ORCID to comment.