paper-with-me

Papers

MMGCN: Multi-modal Graph Convolution Network for Personalized Recommendation of Micro-video

2019-10-19 · ACM International Conference on Multimedia 2019 10 · Yinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He, Richang Hong, Tat-Seng Chua

Personalized recommendation plays a central role in many online content sharing platforms. To provide quality micro-video recommendation service, it is of crucial importance to consider the interactions between users and items (i.e. micro-videos) as well as the item contents from various modalities (e.g. visual, acoustic, and textual). Existing works on multimedia recommendation largely exploit multi-modal contents to enrich item representations, while less effort is made to leverage information interchange between users and items to enhance user representations and further capture user's fine-grained preferences on different modalities. In this paper, we propose to exploit user-item interactions to guide the representation learning in each modality, and further personalized micro-video recommendation. We design a Multi-modal Graph Convolution Network (MMGCN) framework built upon the message-passing idea of graph neural networks, which can yield modal-specific representations of users and micro-videos to better capture user preferences. Specifically, we construct a user-item bipartite graph in each modality, and enrich the representation of each node with the topological structure and features of its neighbors. Through extensive experiments on three publicly available datasets, Tiktok, Kwai, and MovieLens, we demonstrate that our proposed model is able to significantly outperform state-of-the-art multi-modal recommendation methods.

📄 PDF Abstract BibTeX

Code (1)

weiyinwei/mmgcn pytorch

Tasks

Microvideo RecommendationMicro-video recommendationsMultimedia recommendationMulti-Media RecommendationMulti-modal RecommendationMultimodal RecommendationRepresentation Learning

Similar Papers 제목 키워드 기반

MMGCN: Multimodal Fusion via Deep Graph Convolution Network for Emotion Recognition in Conversation

2021-07-14 · ACL 2021 5 · Jingwen Hu, Yuchen Liu, Jinming Zhao, Qin Jin

Emotion recognition in conversation (ERC) is a crucial component in affective dialogue systems, which helps the system understand users' emotions and generate empathetic responses. However, most works focus on modeling s…

Emotion RecognitionEmotion Recognition in Conversation

Multi-Modal Graph Convolutional Network with Sinusoidal Encoding for Robust Human Action Segmentation

2025-07-01 · Hao Xing, Kai Zhe Boey, Yuankai Wu, Darius Burschka 외 arxiv

Accurate temporal segmentation of human actions is critical for intelligent robots in collaborative settings, where a precise understanding of sub-activity labels and their temporal structure is essential. However, the i…

Action SegmentationData AugmentationObject DetectionPose Estimation

GraphCFC: A Directed Graph Based Cross-Modal Feature Complementation Approach for Multimodal Conversational Emotion Recognition

2022-07-06 · Jiang Li, XiaoPing Wang, Guoqing Lv, Zhigang Zeng

Emotion Recognition in Conversation (ERC) plays a significant part in Human-Computer Interaction (HCI) systems since it can provide empathetic services. Multimodal ERC can mitigate the drawbacks of uni-modal approaches. …

Emotion ClassificationEmotion RecognitionEmotion Recognition in ConversationMultimodal Emotion Recognition

Don't Lose Yourself: Boosting Multimodal Recommendation via Reducing Node-neighbor Discrepancy in Graph Convolutional Network

2024-12-25 · Zheyu Chen, Jinfeng Xu, Haibo Hu

The rapid expansion of multimedia contents has led to the emergence of multimodal recommendation systems. It has attracted increasing attention in recommendation systems because its full utilization of data from differen…

Multimodal RecommendationRecommendation Systems

MEGCF: Multimodal Entity Graph Collaborative Filtering for Personalized Recommendation

2022-10-14 · Kang Liu, Feng Xue, Dan Guo, Le Wu 외

In most E-commerce platforms, whether the displayed items trigger the user's interest largely depends on their most eye-catching multimodal content. Consequently, increasing efforts focus on modeling multimodal user pref…

Collaborative Filteringimage-classificationImage Classification