paper-with-me

Papers

The complementarity of a diverse range of deep learning features extracted from video content for video recommendation

2020-11-21 · Adolfo Almeida, Johan Pieter de Villiers, Allan De Freitas, Mergandran Velayudan

Following the popularisation of media streaming, a number of video streaming services are continuously buying new video content to mine the potential profit from them. As such, the newly added content has to be handled well to be recommended to suitable users. In this paper, we address the new item cold-start problem by exploring the potential of various deep learning features to provide video recommendations. The deep learning features investigated include features that capture the visual-appearance, audio and motion information from video content. We also explore different fusion methods to evaluate how well these feature modalities can be combined to fully exploit the complementary information captured by them. Experiments on a real-world video dataset for movie recommendations show that deep learning features outperform hand-crafted features. In particular, recommendations generated with deep learning audio features and action-centric deep learning features are superior to MFCC and state-of-the-art iDT features. In addition, the combination of various deep learning features with hand-crafted features and textual metadata yields significant improvement in recommendations compared to combining only the former.

📄 PDF Abstract BibTeX arXiv:2011.10834

Code (1)

adolfo-almeida/scaled_cer 공식 구현 tf

Tasks

Collaborative FilteringDeep LearningRecommendation SystemsRecommendation Systems (Item cold-start)

Similar Papers 제목 키워드 기반

Two-stream Collaborative Learning with Spatial-Temporal Attention for Video Classification

2017-11-09 · Yuxin Peng, Yunzhen Zhao, Junchao Zhang

Video classification is highly important with wide applications, such as video search and intelligent surveillance. Video naturally consists of static and motion information, which can be represented by frame and optical…

General ClassificationOptical Flow EstimationVideo ClassificationVocal Bursts Valence Prediction

Audio-Visual Dataset and Method for Anomaly Detection in Traffic Videos

2023-05-24 · Błażej Leporowski, Arian Bakhtiarnia, Nicole Bonnici, Adrian Muscat 외

We introduce the first audio-visual dataset for traffic anomaly detection taken from real-world scenes, called MAVAD, with a diverse range of weather and illumination conditions. In addition, we propose a novel method na…

Anomaly Detection

TNTC: two-stream network with transformer-based complementarity for gait-based emotion recognition

2021-10-26 · Chuanfei Hu, Weijie Sheng, Bo Dong, Xinde Li

Recognizing the human emotion automatically from visual characteristics plays a vital role in many intelligent applications. Recently, gait-based emotion recognition, especially gait skeletons-based characteristic, has a…

Emotion Recognition

Diversity-aware Multi-Video Summarization

2017-06-09 · Rameswar Panda, Niluthpol Chowdhury Mithun, Amit K. Roy-Chowdhury

Most video summarization approaches have focused on extracting a summary from a single video; we propose an unsupervised framework for summarizing a collection of videos. We observe that each video in the collection may …

DiversityVideo Summarization

ECLIPSE: Efficient Long-range Video Retrieval using Sight and Sound

2022-04-06 · Yan-Bo Lin, Jie Lei, Mohit Bansal, Gedas Bertasius

We introduce an audiovisual method for long-range text-to-video retrieval. Unlike previous approaches designed for short video retrieval (e.g., 5-15 seconds in duration), our approach aims to retrieve minute-long videos …

RetrievalText to Video RetrievalVideo Retrieval