跳到主要导航 跳到搜索 跳到主要内容

Attention-Based Modality-Gated Networks for Image-Text Sentiment Analysis

  • Jinan University

科研成果: 期刊稿件文章同行评审

摘要

Sentiment analysis of social multimedia data has attracted extensive research interest and has been applied to many tasks, such as election prediction and products evaluation. Sentiment analysis of one modality (e.g., text or image) has been broadly studied. However, not much attention has been paid to the sentiment analysis of multimodal data. Different modalities usually have information that is complementary. Thus, it is necessary to learn the overall sentiment by combining the visual content with text description. In this article, we propose a novel method-Attention-Based Modality-Gated Networks (AMGN)-to exploit the correlation between the modalities of images and texts and extract the discriminative features for multimodal sentiment analysis. Specifically, a visual-semantic attention model is proposed to learn attended visual features for each word. To effectively combine the sentiment information on the two modalities of image and text, a modality-gated LSTM is proposed to learn the multimodal features by adaptively selecting the modality that presents stronger sentiment information. Then a semantic self-attention model is proposed to automatically focus on the discriminative features for sentiment classification. Extensive experiments have been conducted on both manually annotated and machine weakly labeled datasets. The results demonstrate the superiority of our approach through comparison with state-of-the-art models.

源语言英语
文章编号3388861
期刊ACM Transactions on Multimedia Computing, Communications and Applications
16
3
DOI
出版状态已出版 - 9月 2020

指纹

探究 'Attention-Based Modality-Gated Networks for Image-Text Sentiment Analysis' 的科研主题。它们共同构成独一无二的指纹。

引用此