AvaMERG is a new text-speech-vision avatar benchmark for empathetic response generation, and the Empatheia system is claimed to outperform baselines on both textual and multimodal empathy tasks.
CASE: Aligning Coarse-to-Fine Cognition and Affection for Empathetic Response Generation
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Empathetic conversation is psychologically supposed to be the result of conscious alignment and interaction between the cognition and affection of empathy. However, existing empathetic dialogue models usually consider only the affective aspect or treat cognition and affection in isolation, which limits the capability of empathetic response generation. In this work, we propose the CASE model for empathetic dialogue generation. It first builds upon a commonsense cognition graph and an emotional concept graph and then aligns the user's cognition and affection at both the coarse-grained and fine-grained levels. Through automatic and manual evaluation, we demonstrate that CASE outperforms state-of-the-art baselines of empathetic dialogues and can generate more empathetic and informative responses.
citation-role summary
citation-polarity summary
fields
cs.MM 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark
AvaMERG is a new text-speech-vision avatar benchmark for empathetic response generation, and the Empatheia system is claimed to outperform baselines on both textual and multimodal empathy tasks.