Adaptive Graph Learning with Multimodal Fusion for Emotion Recognition in Conversation

Jian Liu; Jian Li; Jiawei Dong; Zifan Mo; Na Liu; Qingdu Li; Ye Yuan

PMC · DOI:10.3390/biomimetics10070414·June 25, 2025

Adaptive Graph Learning with Multimodal Fusion for Emotion Recognition in Conversation

Jian Liu, Jian Li, Jiawei Dong, Zifan Mo, Na Liu, Qingdu Li, Ye Yuan

PDF

Open Access

TL;DR

This paper introduces GASMER, a new model that improves emotion recognition in conversations by combining adaptive graph learning with multimodal data.

Contribution

The novelty lies in the unified architecture that adaptively learns graph structures for modeling conversation dependencies while fusing multimodal data.

Findings

01

GASMER outperforms existing graph-based approaches in emotion recognition.

02

It achieves a 2.7% accuracy improvement on IEMOCAP and 1.2% on MOSEI.

03

The model remains competitive against recent multimodal fusion models.

Abstract

Robust emotion recognition is a prerequisite for natural, fluid human–computer interaction, yet conversational settings remain challenging because emotions are shaped simultaneously by global topic flow and local speaker-to-speaker dependencies. Here, we introduce GASMER—Graph-Adaptive Structure for Multimodal Emotion Recognition—a unified architecture that tackles both issues. It uses the correlation structure based on graph neural networks (GNNs) to model the complex dependencies in the conversation, while adaptively learning the graph structure for GNNs. The experiments indicate that our model has strong performance that outperforms all existing graph-based approaches, and remains competitive when compared to recent multimodal fusion models, underscoring the importance of combining fine-grained multimodal fusion with adaptive graph learning for conversational emotion recognition. On…

Linked entities

Genes, proteins, chemicals, diseases, species, mutations and cell lines named across the full text — each resolved to its canonical identifier and authoritative record.

Species1

Homo sapiens(human · species)

Figures4

Click any figure to enlarge with its caption.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsEmotion and Mood Recognition · Sentiment Analysis and Opinion Mining · Music and Audio Processing