REVIEW 6 cited by
DialogueGCN: A Graph Convolutional Neural Network for Emotion Recognition in Conversation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Emotion recognition in conversation (ERC) has received much attention, lately, from researchers due to its potential widespread applications in diverse areas, such as health-care, education, and human resources. In this paper, we present Dialogue Graph Convolutional Network (DialogueGCN), a graph neural network based approach to ERC. We leverage self and inter-speaker dependency of the interlocutors to model conversational context for emotion recognition. Through the graph network, DialogueGCN addresses context propagation issues present in the current RNN-based methods. We empirically show that this method alleviates such issues, while outperforming the current state of the art on a number of benchmark emotion classification datasets.
Forward citations
Cited by 6 Pith papers
-
TTS-CtrlNet: Time varying emotion aligned text-to-speech generation with ControlNet
TTS-CtrlNet adds time-varying emotion control to a frozen flow-matching TTS model using a ControlNet-style trainable copy, improving emotion similarity metrics while preserving the base model's voice cloning.
-
EmoEUS: Uncertainty Supervision for Multimodal Emotion Recognition in Conversation
Modeling each modality as a Gaussian and supervising its variance with the 2-Wasserstein distance to emotion cluster centers improves IEMOCAP/MELD accuracy by about 0.5-0.8 points over listed baselines.
-
EII-SCL: Harnessing Emotional Inertia for Multimodal Emotion Recognition in Conversation
A plug-in contrastive loss using speaker-local 'emotional inertia' hard negatives improves multimodal emotion-recognition accuracy and F1 on IEMOCAP and MELD by about 0.5–2 points.
-
RAMer: Reconstruction-based Adversarial Model for Multi-party Multi-modal Multi-label Emotion Recognition
RAMer achieves state-of-the-art multi-label emotion recognition on three benchmarks by combining reconstruction-based adversarial training, contrastive learning, a personality cue, and a stack shuffle augmentation to ...
-
Sync-TVA: A Graph-Attention Framework for Multimodal Emotion Recognition with Cross-Modal Fusion
Sync-TVA reports modest accuracy and weighted-F1 improvements over prior graph-based models on MELD and IEMOCAP, using modality-specific enhancement and cross-modal graph fusion.
-
A Survey: Learning Embodied Intelligence from Physical Simulators and World Models
Embodied intelligence learning is reviewed through the complementary lenses of physical simulators and world models, with a proposed IR-L0 to IR-L4 robot capability taxonomy.
Discussion (0). Continue with ORCID to comment.