Video Instance Segmentation
8개 벤치마크 · 논문 163편 · 이 태스크의 논문 보기 →
Benchmarks
OVIS validation
YouTube-VIS validation
YouTube-VIS 2021
BDD100K val
HQ-YTVIS
YouTube-VIS
Most implemented
Simple Online and Realtime Tracking with a Deep Association Metric
Mask2Former for Video Instance Segmentation
Video Instance Segmentation
Instances as Queries
UAV-OVVIS: Unmanned Aerial Vehicles Also Need Open-Vocabulary Video Instance Segmentation
DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries
Papers
UAV-OVVIS: Unmanned Aerial Vehicles Also Need Open-Vocabulary Video Instance Segmentation
Unmanned Aerial Vehicle (UAV) videos are widely used in traffic monitoring, urban management, and emergency rescue. However, existing UAV video perception is largely limited to box-level detection and tracking over prede…
Video Instance SegmentationSegmenting, Fast and Slow: Real-Time Open-Vocabulary Video Instance Segmentation with Dual-Path Processing
Object-centric models inspired by DETR have become the dominant paradigm for open-vocabulary video instance segmentation (OV-VIS). While recent efforts have reduced the computational cost of pixel decoding, textual modal…
Video Instance SegmentationSA-VIS: Sparse frame Annotations for training Video Instance Segmentation
Recent online video instance segmentation (VIS) methods have achieved impressive results, thus becoming the preferred approach to segment instances in videos. Despite the resurgence of impressive single image models, the…
Video Instance SegmentationMind the Gap: Disentangling Performance Bottlenecks in Video Instance Segmentation
In Video Instance Segmentation (VIS), classification, segmentation, and tracking objectives are jointly evaluated, but their individual contributions to performance loss remain opaque. We introduce a diagnostic framework…
Video Instance SegmentationVideo Patch Pruning: Efficient Video Instance Segmentation via Early Token Reduction
Vision Transformers (ViTs) have demonstrated state-ofthe-art performance in several benchmarks, yet their high computational costs hinders their practical deployment. Patch Pruning offers significant savings, but existin…
Video Instance SegmentationSAMannot: A Memory-Efficient, Local, Open-source Framework for Interactive Video Instance Segmentation based on SAM2
Current research workflows for precise video segmentation are often forced into a compromise between labor-intensive manual curation, costly commercial platforms, and/or privacy-compromising cloud-based services. The dem…
Video Instance SegmentationVideo Segmentation