Interactive Multimodal Attention Network for Emotion Recognition in Conversation

Minjie Ren,Xiangdong Huang,Xiaoqi Shi,Weizhi Nie

doi:10.1109/lsp.2021.3078698

Abstract

In this letter, we propose a novel Interactive Multimodal Attention Network (IMAN) for emotion recognition in conversations. IMAN introduces a cross-modal attention fusion module to capture cross-modal interactions of multimodal information, and employs a conversational modeling module to explore the context information and speaker dependency of the whole conversation. Concretely, the cross-modal attention fusion module captures the cross-modal interactions and complementary information among the pre-extracted unimodal features from textual, visual, acoustic modalities based on the cross-modal attention block. Afterward, the updated features from each modality are fused to concentrate more on the informative modality and achieve a refined feature for each constituent utterance. The conversational modeling module defines three different gated recurrent units (GRUs) with respect to the context information, the speaker dependency, and the emotional state of utterances. In this way, we exploit the speaker dependency and contextual information to obtain the emotional state of utterances for emotion classification. Empirical evaluations on the multimodal benchmark IEMOCAP dataset demonstrate that our IMAN achieves competitive performance compared to the state-of-the-art approaches.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Interactive Multimodal Attention Network for Emotion Recognition in Conversation

Abstract

Talk to us

Similar Papers

More From: IEEE Signal Processing Letters

Lead the way for us

Journal: IEEE Signal Processing Letters	Publication Date: Jan 1, 2021
Citations: 15

Similar Papers

HiMul-LGG: A hierarchical decision fusion-based local–global graph neural network for multimodal emotion recognition in conversation
Changzeng Fu ... Carlos Toshinori Ishi
Neural Networks | VOL. 181
Changzeng Fu, et. al.Changzeng Fu ... Carlos Toshinori Ishi
28 Sep 2024
Neural Networks | VOL. 181

Domain Adversarial Network for Cross-Domain Emotion Recognition in Conversation
Hongchao Ma ... Qinglei Zhou
Applied Sciences | VOL. 12
Hongchao Ma, et. al.Hongchao Ma ... Qinglei Zhou
27 May 2022
Applied Sciences | VOL. 12

Transformer-Based Potential Emotional Relation Mining Network for Emotion Recognition in Conversation
Yunwei Shi ... Xiao Sun
-
Yunwei Shi, et. al.Yunwei Shi ... Xiao Sun
01 Jan 2023
01 Jan 2023

M2FNet: Multi-modal Fusion Network for Emotion Recognition in Conversation
Vishal Chudasama ... Naoyuki Onoe
-
Vishal Chudasama, et. al.Vishal Chudasama ... Naoyuki Onoe
01 Jun 2022
01 Jun 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Interactive Multimodal Attention Network for Emotion Recognition in Conversation

Abstract

Talk to us

Similar Papers

More From: IEEE Signal Processing Letters