Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning
Designing an effective representation learning method for multimodal sentiment analysis tasks is a crucial research direction. The challenge lies in learning both shared and private information in a complete modal representation, which is difficult with uniform multimodal labels and a raw feature fusion approach. In this work, we propose a deep modal shared information learning module based on the covariance matrix to capture the shared information between modalities. Additionally, we use a label generation module based on a self-supervised learning strategy to capture the private information of the modalities. Our module is plug-and-play in multimodal tasks, and by changing the parameterization, it can adjust the information exchange relationship between the modes and learn the private or shared information between the specified modes. We also employ a multi-task learning strategy to help the model focus its attention on the modal differentiation training data. We provide a detailed formulation derivation and feasibility proof for the design of the deep modal shared information learning module. We conduct extensive experiments on three common multimodal sentiment analysis baseline datasets, and the experimental results validate the reliability of our model. Furthermore, we explore more combinatorial techniques for the use of the module. Our approach outperforms current state-of-the-art methods on most of the metrics of the three public datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Multimodal Sentiment AnalysisMulti-Task LearningRepresentation LearningSelf-Supervised LearningSentiment AnalysisSimilar Papers 제목 키워드 기반
A Text-Centered Shared-Private Framework via Cross-Modal Prediction for Multimodal Sentiment Analysis
Tri-Subspaces Disentanglement for Multimodal Sentiment Analysis
Multimodal Sentiment Analysis (MSA) integrates language, visual, and acoustic modalities to infer human sentiment. Most existing methods either focus on globally shared representations or modality-specific features, whil…
Multimodal Intent RecognitionMultimodal Sentiment AnalysisCLCR: Cross-Level Semantic Collaborative Representation for Multimodal Learning
Multimodal learning aims to capture both shared and private information from multiple modalities. However, existing methods that project all modalities into a single latent space for fusion often overlook the asynchronou…
Emotion RecognitionSentiment AnalysisAction RecognitionA Shared-Private Representation Model with Coarse-to-Fine Extraction for Target Sentiment Analysis
Target sentiment analysis aims to detect opinion targets along with recognizing their sentiment polarities from a sentence. Some models with span-based labeling have achieved promising results in this task. However, the …
SentenceSentiment AnalysisFindings of the Shared Task on Multimodal Sentiment Analysis and Troll Meme Classification in Dravidian Languages
This paper presents the findings of the shared task on Multimodal Sentiment Analysis and Troll meme classification in Dravidian languages held at ACL 2022. Multimodal sentiment analysis deals with the identification of s…
ClassificationMeme ClassificationMultimodal Sentiment AnalysisSentiment Analysis