REVIEW 2 cited by
Harnessing Large Language Models' Empathetic Response Generation Capabilities for Online Mental Health Counselling Support
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Large Language Models (LLMs) have demonstrated remarkable performance across various information-seeking and reasoning tasks. These computational systems drive state-of-the-art dialogue systems, such as ChatGPT and Bard. They also carry substantial promise in meeting the growing demands of mental health care, albeit relatively unexplored. As such, this study sought to examine LLMs' capability to generate empathetic responses in conversations that emulate those in a mental health counselling setting. We selected five LLMs: version 3.5 and version 4 of the Generative Pre-training (GPT), Vicuna FastChat-T5, Pathways Language Model (PaLM) version 2, and Falcon-7B-Instruct. Based on a simple instructional prompt, these models responded to utterances derived from the EmpatheticDialogues (ED) dataset. Using three empathy-related metrics, we compared their responses to those from traditional response generation dialogue systems, which were fine-tuned on the ED dataset, along with human-generated responses. Notably, we discovered that responses from the LLMs were remarkably more empathetic in most scenarios. We position our findings in light of catapulting advancements in creating empathetic conversational systems.
Forward citations
Cited by 2 Pith papers
-
Exploring User Security and Privacy Attitudes and Concerns Toward the Use of General-Purpose LLM Chatbots for Mental Health
Many users of general-purpose LLM chatbots for mental health misunderstand how their data is protected, conflating human-like empathy with accountability and undervaluing emotional disclosures as a privacy risk.
-
Mentalic Net: Development of RAG-based Conversational AI and Evaluation Framework for Mental Health Support
A RAG-based mental health chatbot using TinyLlama reports a BERT score of 0.898 on a self-created 500-sample evaluation set, with methodological gaps in evaluation design.
Discussion (0). Sign in to comment.