paper-with-me

홈 › Papers

SoccerMaster: A Vision Foundation Model for Soccer Understanding

2025-12-11 · Haolin Yang, Jiayuan Rao, Haoning Wu, Weidi Xie arxiv

Soccer understanding has recently garnered growing research interest due to its domain-specific complexity and unique challenges. Unlike prior works that typically rely on isolated, task-specific expert models, this work aims to propose a unified model to handle diverse soccer visual understanding tasks, ranging from fine-grained perception (e.g., athlete detection and identification) to high-level semantic reasoning (e.g., event classification). Concretely, our contributions are threefold: (i) we present SoccerMaster, the first soccer-specific vision foundation model that unifies diverse tasks within a single framework via supervised multi-task pretraining; (ii) we develop an automated data curation pipeline, SoccerFactory, to generate scalable spatial annotations, and integrate multiple existing soccer video datasets as a comprehensive pretraining data resource for multi-task pretraining; and (iii) we conduct extensive evaluations demonstrating that SoccerMaster consistently outperforms task-specific expert models across diverse downstream tasks, highlighting its breadth and superiority. The data, code, and model will be publicly available.

📄 PDF Abstract BibTeX arXiv:2512.11016

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Universal Soccer Video Understanding

2024-12-02 · CVPR 2025 1 · Jiayuan Rao, HaoNing Wu, Hao Jiang, Ya zhang 외

As a globally celebrated sport, soccer has attracted widespread interest from fans all over the world. This paper aims to develop a comprehensive multi-modal framework for soccer video understanding. Specifically, we mak…

Action ClassificationSports UnderstandingVideo Understanding

AutoSoccerPose: Automated 3D posture Analysis of Soccer Shot Movements

2024-05-20 · Calvin Yeung, Kenjiro Ide, Keisuke Fujii

Image understanding is a foundational task in computer vision, with recent applications emerging in soccer posture analysis. However, existing publicly available datasets lack comprehensive information, notably in the fo…

3D Pose EstimationPose Estimation

SoccerNet-v2: A Dataset and Benchmarks for Holistic Understanding of Broadcast Soccer Videos

2020-11-26 · Adrien Deliège, Anthony Cioppa, Silvio Giancola, Meisam J. Seikavandi 외

Understanding broadcast videos is a challenging task in computer vision, as it requires generic reasoning capabilities to appreciate the content offered by the video editing. In this work, we propose SoccerNet-v2, a nove…

Action SpottingBoundary DetectionCamera shot boundary detectionCamera shot segmentation+3

SoccerNet 2024 Challenges Results

2024-09-16 · Anthony Cioppa, Silvio Giancola, Vladimir Somers, Victor Joos 외

The SoccerNet 2024 challenges represent the fourth annual video understanding challenges organized by the SoccerNet team. These challenges aim to advance research across multiple themes in football, including broadcast v…

Action SpottingDense Video CaptioningGame State ReconstructionVideo Captioning+1

Domain Adaptation of VLM for Soccer Video Understanding

2025-05-20 · Tiancheng Jiang, Henry Wang, Md Sirajus Salekin, Parmida Atighehchian 외

Vision Language Models (VLMs) have demonstrated strong performance in multi-modal tasks by effectively aligning visual and textual representations. However, most video understanding VLM research has been domain-agnostic,…

Action ClassificationDomain AdaptationInstruction FollowingQuestion Answering+3