Papers › Multivariate, Multi-Frequency and Multimodal: Rethinking Graph Neural Networks for...

Multivariate, Multi-Frequency and Multimodal: Rethinking Graph Neural Networks for Emotion Recognition in Conversation

1 Jan 2023CVPR 2023 1archive 2025-07-28

Feiyu Chen, Jie Shao, Shuyuan Zhu, Heng Tao Shen

Complex relationships of high arity across modality and context dimensions is a critical challenge in the Emotion Recognition in Conversation (ERC) task. Yet, previous works tend to encode multimodal and contextual relationships in a loosely-coupled manner, which may harm relationship modelling. Recently, Graph Neural Networks (GNN) which show advantages in capturing data relations, offer a new solution for ERC. However, existing GNN-based ERC models fail to address some general limits of GNNs, including assuming pairwise formulation and erasing high-frequency signals, which may be trivial for many applications but crucial for the ERC task. In this paper, we propose a GNN-based model that explores multivariate relationships and captures the varying importance of emotion discrepancy and commonality by valuing multi-frequency signals. We empower GNNs to better capture the inherent relationships among utterances and deliver more sufficient multimodal and contextual modelling. Experimental results show that our proposed method outperforms previous state-of-the-art works on two popular multimodal ERC datasets.

PaperPDFCode

Code

feiyuchen7/M3NET officialpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Emotion RecognitionEmotion Recognition in Conversation

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Emotion Recognition in Conversation CMU-MOSEI-Sentiment M3Net Accuracy 43.67 #5 of 7 Archive leaderboard report
Emotion Recognition in Conversation CMU-MOSEI-Sentiment M3Net Weighted F1 41.12 #5 of 7 Archive leaderboard report
Emotion Recognition in Conversation IEMOCAP-4 M3Net Accuracy 83.67 #4 of 8 Archive leaderboard report
Emotion Recognition in Conversation IEMOCAP-4 M3Net Weighted F1 83.57 #4 of 8 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

fail

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections