LLM safety guardrails fail for most mental health conditions with up to 100% failure rates for eating disorders, substance use disorder, and major depressive disorder, while holding only for suicide and self-harm.
Large language models in mental health care: a scoping review
3 Pith papers cite this work. Polarity classification is still indexing.
years
2026 3verdicts
UNVERDICTED 3representative citing papers
DySRec is a multi-agent conversational system that dynamically recommends psychometric scales by integrating user context, behaviors, and risk signals through interactive dialogue and closed-loop refinement.
Topic-aware augmentation makes psychosocial risk factors such as immigration, family issues, and financial crisis more distinct and coherent in the internal representations of suicide ideation detection models.
citing papers explorer
-
One Year Later...The Harms Persist, But So Do We!
LLM safety guardrails fail for most mental health conditions with up to 100% failure rates for eating disorders, substance use disorder, and major depressive disorder, while holding only for suicide and self-harm.
-
DySRec: Dynamic Context-Aware Psychometric Scale Recommendation via Multi-Agent Collaboration
DySRec is a multi-agent conversational system that dynamically recommends psychometric scales by integrating user context, behaviors, and risk signals through interactive dialogue and closed-loop refinement.
-
Beyond Accuracy: Interpreting Topic Representation in Suicide Ideation Detection Models
Topic-aware augmentation makes psychosocial risk factors such as immigration, family issues, and financial crisis more distinct and coherent in the internal representations of suicide ideation detection models.