paper-with-me

Papers

Multimodal Topic Learning for Video Recommendation

2020-10-26 · Shi Pu, Yijiang He, Zheng Li, Mao Zheng

Facilitated by deep neural networks, video recommendation systems have made significant advances. Existing video recommendation systems directly exploit features from different modalities (e.g., user personal data, user behavior data, video titles, video tags, and visual contents) to input deep neural networks, while expecting the networks to online mine user-preferred topics implicitly from these features. However, the features lacking semantic topic information limits accurate recommendation generation. In addition, feature crosses using visual content features generate high dimensionality features that heavily downgrade the online computational efficiency of networks. In this paper, we explicitly separate topic generation from recommendation generation, propose a multimodal topic learning algorithm to exploit three modalities (i.e., tags, titles, and cover images) for generating video topics offline. The topics generated by the proposed algorithm serve as semantic topic features to facilitate preference scope determination and recommendation generation. Furthermore, we use the semantic topic features instead of visual content features to effectively reduce online computational cost. Our proposed algorithm has been deployed in the Kuaibao information streaming platform. Online and offline evaluation results show that our proposed algorithm performs favorably.

📄 PDF Abstract BibTeX arXiv:2010.13373

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyRecommendation Systems

Similar Papers 제목 키워드 기반

FinCap: Topic-Aligned Captions for Short-Form Financial YouTube Videos

2025-09-30 · Siddhant Sukhani, Yash Bhardwaj, Riya Bhadani, Veer Kejriwal 외 arxiv

We evaluate multimodal large language models (MLLMs) for topic-aligned captioning in financial short-form videos (SVs) by testing joint reasoning over transcripts (T), audio (A), and video (V). Using 624 annotated YouTub…

Sentiment AnalysisVideo Captioning

Research on the Design of a Short Video Recommendation System Based on Multimodal Information and Differential Privacy

2025-03-27 · Haowei Yang, Lei Fu, Qingyi Lu, Yue Fan 외

With the rapid development of short video platforms, recommendation systems have become key technologies for improving user experience and enhancing platform engagement. However, while short video recommendation systems …

Recommendation Systems

A Multimodal Sentiment Dataset for Video Recommendation

2021-09-17 · Hongxuan Tang, Hao liu, Xinyan Xiao, Hua Wu

Recently, multimodal sentiment analysis has seen remarkable advance and a lot of datasets are proposed for its development. In general, current multimodal sentiment analysis datasets usually follow the traditional system…

Multimodal Sentiment AnalysisSentiment AnalysisVideo Understanding

Video Captioning with Guidance of Multimodal Latent Topics

2017-08-31 · Shizhe Chen, Jia Chen, Qin Jin, Alexander Hauptmann

The topic diversity of open-domain videos leads to various vocabularies and linguistic expressions in describing video contents, and therefore, makes the video captioning task even more challenging. In this paper, we pro…

Caption GenerationDecoderMulti-Task LearningPrediction+1

Dynamic Multimodal Fusion via Meta-Learning Towards Micro-Video Recommendation

2025-01-13 · Han Liu, Yinwei Wei, Fan Liu, Wenjie Wang 외

Multimodal information (e.g., visual, acoustic, and textual) has been widely used to enhance representation learning for micro-video recommendation. For integrating multimodal information into a joint representation of m…

Meta-LearningMultimodal RecommendationRepresentation Learning