paper-with-me

홈 › Papers

MVP: Winning Solution to SMP Challenge 2025 Video Track

2025-07-01 · Liliang Ye, Yunyao Zhang, Yafeng Wu, Yi-Ping Phoebe Chen, Junqing Yu, Wei Yang, Zikai Song arxiv

Social media platforms serve as central hubs for content dissemination, opinion expression, and public engagement across diverse modalities. Accurately predicting the popularity of social media videos enables valuable applications in content recommendation, trend detection, and audience engagement. In this paper, we present Multimodal Video Predictor (MVP), our winning solution to the Video Track of the SMP Challenge 2025. MVP constructs expressive post representations by integrating deep video features extracted from pretrained models with user metadata and contextual information. The framework applies systematic preprocessing techniques, including log-transformations and outlier removal, to improve model robustness. A gradient-boosted regression model is trained to capture complex patterns across modalities. Our approach ranked first in the official evaluation of the Video Track, demonstrating its effectiveness and reliability for multimodal video popularity prediction on social platforms. The source code is available at https://anonymous.4open.science/r/SMPDVideo.

📄 PDF Abstract BibTeX arXiv:2507.00950

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

NTIRE 2020 Challenge on Image and Video Deblurring

2020-05-04 · Seungjun Nah, Sanghyun Son, Radu Timofte, Kyoung Mu Lee

Motion blur is one of the most common degradation artifacts in dynamic scene photography. This paper reviews the NTIRE 2020 Challenge on Image and Video Deblurring. In this challenge, we present the evaluation results fr…

DeblurringImage DeblurringSingle Image DeblurringVideo Deblurring

The 1st Winner for 5th PVUW MeViS-Text Challenge: Strong MLLMs Meet SAM3 for Referring Video Object Segmentation

2026-04-01 · Xusheng He, Canyang Wu, Jinrong Zhang, Weili Guan 외 arxiv

This report presents our winning solution to the 5th PVUW MeViS-Text Challenge. The track studies referring video object segmentation under motion-centric language expressions, where the model must jointly understand app…

Referring Video Object Segmentation

Detect to Track and Track to Detect

2017-10-11 · ICCV 2017 10 · Christoph Feichtenhofer, Axel Pinz, Andrew Zisserman

Recent approaches for high accuracy detection and tracking of object categories in video consist of complex multistage solutions that become more cumbersome each year. In this paper we propose a ConvNet architecture that…

Objectobject-detectionObject Detection

NTIRE 2024 Quality Assessment of AI-Generated Content Challenge

2024-04-25 · Xiaohong Liu, Xiongkuo Min, Guangtao Zhai, Chunyi Li 외

This paper reports on the NTIRE 2024 Quality Assessment of AI-Generated Content Challenge, which will be held in conjunction with the New Trends in Image Restoration and Enhancement Workshop (NTIRE) at CVPR 2024. This ch…

Image Quality AssessmentImage RestorationVideo Quality AssessmentVisual Question Answering (VQA)

MixMatch Domain Adaptaion: Prize-winning solution for both tracks of VisDA 2019 challenge

2019-10-09 · Danila Rukhovich, Danil Galeev

We present a domain adaptation (DA) system that can be used in multi-source and semi-supervised settings. Using the proposed method we achieved 2nd place on multi-source track and 3rd place on semi-supervised track of th…

Domain Adaptation