Video Panoptic Segmentation
5개 벤치마크 · 논문 44편 · 이 태스크의 논문 보기 →
Benchmarks
Most implemented
A Simple Video Segmenter by Tracking Objects Along Axial Trajectories
ViP-DeepLab: Learning Visual Perception with Depth-aware Video Panoptic Segmentation
MM-OR: A Large Multimodal Operating Room Dataset for Semantic Understanding of High-Intensity Surgical Environments
Context-Aware Video Instance Segmentation
Uni-DVPS: Unified Model for Depth-Aware Video Panoptic Segmentation
Papers
Scene-Centric Unsupervised Video Panoptic Segmentation
Video panoptic segmentation (VPS) aims to jointly detect, segment, and track all objects while partitioning the video into semantically consistent regions. We introduce the task setting of unsupervised VPS, omitting any …
Video Panoptic SegmentationScene UnderstandingVideo SegmentationImage SegmentationSPORTS: Simultaneous Panoptic Odometry, Rendering, Tracking and Segmentation for Urban Scenes Understanding
The scene perception, understanding, and simulation are fundamental techniques for embodied-AI agents, while existing solutions are still prone to segmentation deficiency, dynamic objects' interference, sensor data spars…
Video Panoptic SegmentationCamera Pose EstimationNovel View SynthesisScene UnderstandingA Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects
Video Scene Parsing (VSP) has emerged as a cornerstone in computer vision, facilitating the simultaneous segmentation, recognition, and tracking of diverse visual entities in dynamic scenes. In this survey, we present a …
BenchmarkingInstance SegmentationOpen-Vocabulary Video SegmentationPanoptic Segmentation+8MM-OR: A Large Multimodal Operating Room Dataset for Semantic Understanding of High-Intensity Surgical Environments
Operating rooms (ORs) are complex, high-stakes environments requiring precise understanding of interactions among medical staff, tools, and equipment for enhancing surgical assistance, situational awareness, and patient …
2D Panoptic SegmentationGraph GenerationLanguage ModelingLanguage Modelling+2LiDAR-Camera Fusion for Video Panoptic Segmentation without Video Training
Panoptic segmentation, which combines instance and semantic segmentation, has gained a lot of attention in autonomous vehicles, due to its comprehensive representation of the scene. This task can be applied for cameras a…
Autonomous VehiclesPanoptic SegmentationSegmentationSemantic Segmentation+1Balancing Shared and Task-Specific Representations: A Hybrid Approach to Depth-Aware Video Panoptic Segmentation
In this work, we present Multiformer, a novel approach to depth-aware video panoptic segmentation (DVPS) based on the mask transformer paradigm. Our method learns object representations that are shared across segmentatio…
DecoderDepth-aware Video Panoptic SegmentationDepth EstimationMonocular Depth Estimation+4