Pith. sign in

REVIEW 3 cited by

Large Language Models Produce Responses Perceived to be Empathic

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.18148 v1 pith:2FJTTL5W submitted 2024-03-26 cs.CL cs.AI

classification cs.CLcs.AI
keywords modelsresponsesempathicempathyhumanlanguagelargellms
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Large Language Models (LLMs) have demonstrated surprising performance on many tasks, including writing supportive messages that display empathy. Here, we had these models generate empathic messages in response to posts describing common life experiences, such as workplace situations, parenting, relationships, and other anxiety- and anger-eliciting situations. Across two studies (N=192, 202), we showed human raters a variety of responses written by several models (GPT4 Turbo, Llama2, and Mistral), and had people rate these responses on how empathic they seemed to be. We found that LLM-generated responses were consistently rated as more empathic than human-written responses. Linguistic analyses also show that these models write in distinct, predictable ``styles", in terms of their use of punctuation, emojis, and certain words. These results highlight the potential of using LLMs to enhance human peer support in contexts where empathy is important.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Automating and Scaling Behavioral Scientific Research on AI Agents

    cs.AI 2026-08 conditional novelty 7.0 of 10

    AEROBAT, an LLM-based multi-agent system, automates the full pipeline of behavioral research on AI agents and reports moderate-to-strong evidence for 26 of 79 tested hypotheses.

  2. Generative AI may backfire for counterspeech

    cs.SI 2024-11 conditional novelty 7.0 of 10

    A pre-registered field experiment finds that generic warning counterspeech can reduce hate posting, while LLM-generated contextualized counterspeech is ineffective and may backfire.

  3. The Illusion of Empathy: How AI Chatbots Shape Conversation Perception

    cs.HC 2024-11 conditional novelty 6.0 of 10

    In chat conversations, users rate AI chatbots as less empathetic than humans but still give the chatbot conversations higher quality ratings.

Pith tools