paper-with-me

Papers

PVUW 2025 Challenge Report: Advances in Pixel-level Understanding of Complex Videos in the Wild

2025-04-15 · Henghui Ding, Chang Liu, Nikhila Ravi, Shuting He, Yunchao Wei, Song Bai, Philip Torr, Kehuan Song, Xinglin Xie, Kexin Zhang, Licheng Jiao, Lingling Li, Shuyuan Yang, Xuqiang Cao, Linnan Zhao, Jiaxuan Zhao, Fang Liu, Mengjiao Wang, Junpei Zhang, Xu Liu, Yuting Yang, Mengru Ma, Hao Fang, Runmin Cong, Xiankai Lu, Zhiyang Chen, Wei zhang, Tianming Liang, Haichao Jiang, Wei-Shi Zheng, Jian-Fang Hu, Haobo Yuan, Xiangtai Li, Tao Zhang, Lu Qi, Ming-Hsuan Yang

This report provides a comprehensive overview of the 4th Pixel-level Video Understanding in the Wild (PVUW) Challenge, held in conjunction with CVPR 2025. It summarizes the challenge outcomes, participating methodologies, and future research directions. The challenge features two tracks: MOSE, which focuses on complex scene video object segmentation, and MeViS, which targets motion-guided, language-based video segmentation. Both tracks introduce new, more challenging datasets designed to better reflect real-world scenarios. Through detailed evaluation and analysis, the challenge offers valuable insights into the current state-of-the-art and emerging trends in complex video segmentation. More information can be found on the workshop website: https://pvuw.github.io/.

📄 PDF Abstract BibTeX arXiv:2504.11326

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationSemantic SegmentationVideo Object SegmentationVideo SegmentationVideo Semantic SegmentationVideo Understanding

Similar Papers 제목 키워드 기반

Report of the 5th PVUW Challenge: Towards More Diverse Modalities in Pixel-Level Understanding

2026-04-28 · Chang Liu, Henghui Ding, Nikhila Ravi, Yunchao Wei 외 arxiv

This report summarizes the objectives, datasets, and top-performing methodologies of the 2026 Pixel-level Video Understanding in the Wild (PVUW) Challenge, hosted at CVPR 2026, which evaluates state-of-the-art models und…

Object Segmentation

1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation

2024-06-08 · Qingfeng Liu, Mostafa El-Khamy, Kee-Bong Song

The third Pixel-level Video Understanding in the Wild (PVUW CVPR 2024) challenge aims to advance the state of art in video understanding through benchmarking Video Panoptic Segmentation (VPS) and Video Semantic Segmentat…

BenchmarkingInstance SegmentationPanoptic SegmentationScene Parsing+6

3rd Place of MeViS-Audio Track of the 5th PVUW: VIRST-Audio

2026-03-24 · Jihwan Hong, Jaeyoung Do arxiv

Audio-based Referring Video Object Segmentation (ARVOS) requires grounding audio queries into pixel-level object masks over time, posing challenges in bridging acoustic signals with spatio-temporal visual representations…

Referring Video Object SegmentationVideo Segmentation

PVUW 2024 Challenge on Complex Video Understanding: Methods and Results

2024-06-24 · Henghui Ding, Chang Liu, Yunchao Wei, Nikhila Ravi 외

Pixel-level Video Understanding in the Wild Challenge (PVUW) focus on complex video understanding. In this CVPR 2024 workshop, we add two new tracks, Complex Video Object Segmentation Track based on MOSE dataset and Moti…

SegmentationSemantic SegmentationvalidVideo Object Segmentation+3

MASSeg : 2nd Technical Report for 4th PVUW MOSE Track

2025-04-14 · Xuqiang Cao, Linnan Zhao, Jiaxuan Zhao, Fang Liu 외

Complex video object segmentation continues to face significant challenges in small object recognition, occlusion handling, and dynamic scene modeling. This report presents our solution, which ranked second in the MOSE t…

Data AugmentationObjectObject RecognitionOcclusion Handling+4