REVIEW 15 cited by
Is ChatGPT Equipped with Emotional Dialogue Capabilities?
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
This report presents a study on the emotional dialogue capability of ChatGPT, an advanced language model developed by OpenAI. The study evaluates the performance of ChatGPT on emotional dialogue understanding and generation through a series of experiments on several downstream tasks. Our findings indicate that while ChatGPT's performance on emotional dialogue understanding may still lag behind that of supervised models, it exhibits promising results in generating emotional responses. Furthermore, the study suggests potential avenues for future research directions.
Forward citations
Cited by 15 Pith papers
-
Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness
PsySET measures emotion and personality steering in LLMs across prompting, fine-tuning, and representation engineering, finding prompts most effective overall and emotion-specific safety trade-offs (e.g., joy weakens ...
-
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards
An end-to-end RL framework that uses simulated future dialogue and a learned future-oriented reward model to fine-tune LLMs for open-ended emotional support, reporting improved success rates on ESConv and ExTES.
-
Contrastive Distillation of Emotion Knowledge from LLMs for Zero-Shot Emotion Recognition
A BERT-sized model, trained with contrastive learning on GPT-4-generated emotion descriptors, achieves zero-shot emotion recognition across new label spaces and task types.
-
Simulating Before Planning: Constructing Intrinsic User World Model for User-Tailored Dialogue Policy Planning
UDP, a user-tailored dialogue policy planner with a diffusion-based persona portrayer and a Brownian Bridge feedback anticipator, outperforms existing planners on simulated persuasion and emotional-support tasks.
-
Simulation-Free Hierarchical Latent Policy Planning for Proactive Dialogues
LDPP automatically discovers latent dialogue policies from raw records and uses offline hierarchical reinforcement learning to plan in that latent space, outperforming strong baselines on proactive dialogue benchmarks.
-
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models
IRR restores safety to fine-tuned LLMs by masking delta parameters that conflict with a safety vector, then recalibrating the survivors with inverse-Hessian compensation to preserve task performance.
-
Rethinking Emotion Annotations in the Era of Large Language Models
Human evaluators preferred GPT-4's zero-shot emotion labels over original human labels in 62% of disagreement samples, and GPT-4 pre-filtering and post-filtering can reduce annotation workload and improve training efficiency.
-
MEMO-Bench: A Multiple Benchmark for Text-to-Image and Multimodal Large Language Models on Human Emotion Analysis
MEMO-Bench scores 12 text-to-image models and 16 multimodal LLMs on emotion generation and recognition, finding stronger performance on positive emotions and weak fine-grained intensity estimation.
-
Mechanistic Interpretability of Emotion Inference in Large Language Models
Emotion inference in LLMs is localized to mid-layer attention and feed-forward units, and steering learned appraisal directions shifts generated emotions in appraisal-theory-consistent ways.
-
A Comprehensive Evaluation of Large Language Models on Aspect-Based Sentiment Analysis
Across 13 datasets and 8 ABSA subtasks, efficiently fine-tuned LLMs outperform cited fine-tuned SLM baselines, and retrieval-based demonstration selection improves in-context learning for API models.
-
From Intents to Conversations: Generating Intent-Driven Dialogues with Contrastive Learning for Multi-Turn Classification
An LLM-enhanced HMM generates intent-aware multilingual e-commerce dialogues, and a contrastive multi-task classifier (MINT-CL) improves multi-turn intent classification accuracy by about 0.5 percent on average.
-
Multi-Party Conversational Agents: A Survey
A survey of multi-party conversational AI that organizes tasks into state-of-mind modeling, semantic understanding, and action modeling, and argues that theory of mind is the key missing ingredient.
-
MADP: Multi-Agent Deductive Planning for Enhanced Cognitive-Behavioral Mental Health Question Answer
The MADP framework uses Explorer, Empathizer, and Interpreter agents to plan LLM mental health responses, reporting roughly 4-5% improvements that may be fragile due to weak evaluation.
-
Strategic Prompting for Conversational Tasks: A Comparative Analysis of Large Language Models Across Diverse Conversational Tasks
No single open-source LLM among Llama, OPT, Falcon, Alpaca, and MPT performs best across reservation, empathy, counseling, persuasion, and negotiation tasks.
-
Generalist Virtual Agents: A Survey on Autonomous Agents Across Digital Platforms
A survey that proposes the Generalist Virtual Agent concept and taxonomies for agent environments, tasks, perceptions, actions, models, and evaluation, concluding that real-world-like environments favor human-like int...
Discussion (0). Continue with ORCID to comment.