REVIEW 10 cited by
DialogueGCN: A Graph Convolutional Neural Network for Emotion Recognition in Conversation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Emotion recognition in conversation (ERC) has received much attention, lately, from researchers due to its potential widespread applications in diverse areas, such as health-care, education, and human resources. In this paper, we present Dialogue Graph Convolutional Network (DialogueGCN), a graph neural network based approach to ERC. We leverage self and inter-speaker dependency of the interlocutors to model conversational context for emotion recognition. Through the graph network, DialogueGCN addresses context propagation issues present in the current RNN-based methods. We empirically show that this method alleviates such issues, while outperforming the current state of the art on a number of benchmark emotion classification datasets.
Forward citations
Cited by 10 Pith papers
-
TTS-CtrlNet: Time varying emotion aligned text-to-speech generation with ControlNet
TTS-CtrlNet adds time-varying emotion control to a frozen flow-matching TTS model using a ControlNet-style trainable copy, improving emotion similarity metrics while preserving the base model's voice cloning.
-
EmoEUS: Uncertainty Supervision for Multimodal Emotion Recognition in Conversation
Modeling each modality as a Gaussian and supervising its variance with the 2-Wasserstein distance to emotion cluster centers improves IEMOCAP/MELD accuracy by about 0.5-0.8 points over listed baselines.
-
EII-SCL: Harnessing Emotional Inertia for Multimodal Emotion Recognition in Conversation
A plug-in contrastive loss using speaker-local 'emotional inertia' hard negatives improves multimodal emotion-recognition accuracy and F1 on IEMOCAP and MELD by about 0.5–2 points.
-
RAMer: Reconstruction-based Adversarial Model for Multi-party Multi-modal Multi-label Emotion Recognition
RAMer achieves state-of-the-art multi-label emotion recognition on three benchmarks by combining reconstruction-based adversarial training, contrastive learning, a personality cue, and a stack shuffle augmentation to ...
-
Adaptive Progressive Attention Graph Neural Network for EEG Emotion Recognition
A three-expert progressive attention graph neural network improves EEG emotion classification accuracy on SEED, SEED-IV, and MPED benchmarks.
-
EmoVerse: Exploring Multimodal Large Language Models for Sentiment and Emotion Understanding
A new MLLM and dataset combining five affect tasks with a multi-stage instruction-tuning strategy yields strong results on sentiment and emotion benchmarks, but the empirical setup has unresolved comparison and data-r...
-
Sync-TVA: A Graph-Attention Framework for Multimodal Emotion Recognition with Cross-Modal Fusion
Sync-TVA reports modest accuracy and weighted-F1 improvements over prior graph-based models on MELD and IEMOCAP, using modality-specific enhancement and cross-modal graph fusion.
-
A Survey: Learning Embodied Intelligence from Physical Simulators and World Models
Embodied intelligence learning is reviewed through the complementary lenses of physical simulators and world models, with a proposed IR-L0 to IR-L4 robot capability taxonomy.
-
SpikEmo: Enhancing Emotion Recognition With Spiking Temporal Dynamics in Conversations
SpikEmo stacks spiking self-attention layers on text, audio, and video features with DSC and correlation losses, reporting 65.92 and 71.50 weighted-F1 on MELD and IEMOCAP.
-
Towards Data-centric Machine Learning on Directed Graphs: a Survey
A survey taxonomizing directed graph neural networks into message-passing, eigenpolynomial, and sequence-based frameworks and re-reading them from a data-centric perspective.
Discussion (0). Continue with ORCID to comment.