Dynamic hypergraph convolutional network for multimodal sentiment analysis

超图计算机科学成对比较图形理论计算机科学模态（人机交互）人工智能仿射变换数学离散数学纯数学

作者

Jian Huang,Yuanyuan Pu,Dongming Zhou,Jinde Cao,Jinjing Gu,Zhengpeng Zhao,Dan Xu

出处

期刊：Neurocomputing [Elsevier]
日期：2023-11-02 卷期号：565: 126992-126992 被引量：31

标识

DOI：10.1016/j.neucom.2023.126992

摘要

Multimodal sentiment analysis (MSA) aims to detect the sentiments from language (text), audio, and visual (facial expressions) modalities. The main challenge in MSA is how to efficiently model intra-modality and inter-modality dynamics. With the advent of graph convolution network (GCN), graph-based models are proposed to solve the challenge. However, general graphs contain only two nodes per edge, which limits the exploitation of high-order interactions. Moreover, current graph-based models mainly aggregate the features of each node during fusion, while the features of connected edges are not well mined. In this paper, we introduce dynamic hypergraph convolution networks to MSA for the first time and propose a Multimodal Dynamic Hypergraph Network (MDH) to learn intra- and inter-modality dynamics. Hypergraphs provide a natural approach to capture transcendental pairwise relations, and their potential for MSA remains unexplored. MDH mainly consists of three components: Unimodal Encoder, Dynamic Hypergraph Enhancement Network (DHEN), and HyperFusion module. Specifically, DHEN is composed of Cross-modal Affine, Hypergraph Construction, and Hypergraph Aggregation modules. As for the intra-modality dynamics, MDH utilizes Hypergraph Construction and Aggregation modules to model the interactions within time steps for each modality. As for the inter-modality dynamics, MDH implements Cross-modal Affine and HyperFusion modules to learn the relationships of the modalities. In addition, multi-task learning has been implemented to optimize the learning process for multimodal tasks. Experiments show that MDH outperforms graph-based models on CMU-MOSI and CMU-MOSEI datasets, as well as obtains new state-of-the-art results on CH-SIMS dataset. Furthermore, we conduct external experiments to explore the effectiveness of MDH and the effect of model depth with different graph networks.

求助该文献

最长约 10秒，即可获得该文献文件

Dynamic hypergraph convolutional network for multimodal sentiment analysis

今日热心研友