paper-with-me

Papers

Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization

2025-06-10 · Zhiyi Zhu, Xiaoyu Wu, Youwei Lu

Video memorability refers to the ability of videos to be recalled after viewing, playing a crucial role in creating content that remains memorable. Existing models typically focus on extracting multimodal features to predict video memorability scores but often fail to fully utilize motion cues. The representation of motion features is compromised during the fine-tuning phase of the motion feature extractor due to a lack of labeled data. In this paper, we introduce the Text-Motion Cross-modal Contrastive Loss (TMCCL), a multimodal video memorability prediction model designed to enhance the representation of motion features. We tackle the challenge of improving motion feature representation by leveraging text description similarities across videos to establish positive and negative motion sample sets for a given target. This enhancement allows the model to learn similar feature representations for semantically related motion content, resulting in more accurate memorability predictions. Our model achieves state-of-the-art performance on two video memorability prediction datasets. Moreover, the potential applications of video memorability prediction have been underexplored. To address this gap, we present Memorability Weighted Correction for Video Summarization (MWCVS), using video memorability prediction to reduce subjectivity in video summarization labels. Experimental results on two video summarization datasets demonstrate the effectiveness of MWCVS, showcasing the promising applications of video memorability prediction.

📄 PDF Abstract BibTeX arXiv:2506.08649

Code (0)

등록된 구현이 없습니다.

Tasks

PredictionVideo Summarization

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Memorability: An image-computable measure of information utility

2021-04-01 · Zoya Bylinskii, Lore Goetschalckx, Anelise Newman, Aude Oliva

The pixels in an image, and the objects, scenes, and actions that they compose, determine whether an image will be memorable or forgettable. While memorability varies by image, it is largely independent of an individual …

Prediction

Using Saliency and Cropping to Improve Video Memorability

2023-09-21 · Vaibhav Mudgal, Qingyang Wang, Lorin Sweeney, Alan F. Smeaton

Video memorability is a measure of how likely a particular video is to be remembered by a viewer when that viewer has no emotional connection with the video content. It is an important characteristic as videos that are m…

Position

VideoMem: Constructing, Analyzing, Predicting Short-term and Long-term Video Memorability

2018-12-05 · ICCV 2019 10 · Romain Cohendet, Claire-Hélène Demarty, Ngoc Q. K. Duong, Martin Engilberge

Humans share a strong tendency to memorize/forget some of the visual information they encounter. This paper focuses on providing computational models for the prediction of the intrinsic memorability of visual content. To…

MemorizationPrediction

Overview of MediaEval 2020 Predicting Media Memorability Task: What Makes a Video Memorable?

2020-12-31 · Alba García Seco De Herrera, Rukiye Savran Kiziltepe, Jon Chamberlain, Mihai Gabriel Constantin 외

This paper describes the MediaEval 2020 \textit{Predicting Media Memorability} task. After first being proposed at MediaEval 2018, the Predicting Media Memorability task is in its 3rd edition this year, as the prediction…

Generative Outpainting To Enhance the Memorability of Short-Form Videos

2024-11-21 · Alan Byju, Aman Sudhindra Ladwa, Lorin Sweeney, Alan F. Smeaton

With the expanding use of the short-form video format in advertising, social media, entertainment, education and more, there is a need for such media to both captivate and be remembered. Video memorability indicates to u…

Form