paper-with-me

Papers

Learning Video Object Segmentation from Static Images

2016-12-08 · CVPR 2017 7 · Anna Khoreva, Federico Perazzi, Rodrigo Benenson, Bernt Schiele, Alexander Sorkine-Hornung

Inspired by recent advances of deep learning in instance segmentation and object tracking, we introduce video object segmentation problem as a concept of guided instance segmentation. Our model proceeds on a per-frame basis, guided by the output of the previous frame towards the object of interest in the next frame. We demonstrate that highly accurate object segmentation in videos can be enabled by using a convnet trained with static images only. The key ingredient of our approach is a combination of offline and online learning strategies, where the former serves to produce a refined mask from the previous frame estimate and the latter allows to capture the appearance of the specific object instance. Our method can handle different types of input annotations: bounding boxes and segments, as well as incorporate multiple annotated frames, making the system suitable for diverse applications. We obtain competitive results on three different datasets, independently from the type of input annotation.

📄 PDF Abstract BibTeX arXiv:1612.02646

Code (2)

birdman9390/MetaMaskTrack pytorch
omkar13/MaskTrack pytorch

Tasks

Instance SegmentationObjectObject TrackingSegmentationSemantic SegmentationSemi-Supervised Video Object SegmentationVideo Object SegmentationVideo Semantic SegmentationVisual Object Tracking

Similar Papers 제목 키워드 기반

Instance Embedding Transfer to Unsupervised Video Object Segmentation

2018-01-03 · CVPR 2018 6 · Siyang Li, Bryan Seybold, Alexey Vorobyov, Alireza Fathi 외

We propose a method for unsupervised video object segmentation by transferring the knowledge encapsulated in image-based instance embedding networks. The instance embedding network produces an embedding vector for each p…

ObjectOptical Flow EstimationSegmentationSemantic Segmentation+3

DyStaB: Unsupervised Object Segmentation via Dynamic-Static Bootstrapping

2020-08-16 · CVPR 2021 1 · Yanchao Yang, Brian Lai, Stefano Soatto

We describe an unsupervised method to detect and segment portions of images of live scenes that, at some point in time, are seen moving as a coherent whole, which we refer to as objects. Our method first partitions the m…

Continual LearningObjectobject-detectionObject Detection+7

HODOR: High-level Object Descriptors for Object Re-segmentation in Video Learned from Static Images

2021-12-16 · CVPR 2022 1 · Ali Athar, Jonathon Luiten, Alexander Hermans, Deva Ramanan 외

Existing state-of-the-art methods for Video Object Segmentation (VOS) learn low-level pixel-to-pixel correspondences between frames to propagate object masks across video. This requires a large amount of densely annotate…

ObjectSemantic SegmentationVideo Object SegmentationVideo Semantic Segmentation

Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation

2024-04-21 · Gensheng Pei, Yazhou Yao, Jianbo Jiao, Wenguan Wang 외

Conventional video object segmentation (VOS) methods usually necessitate a substantial volume of pixel-level annotated video data for fully supervised learning. In this paper, we present HVC, a \textbf{h}ybrid static-dyn…

Semantic SegmentationVideo Object SegmentationVideo Semantic Segmentation

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

2025-01-07 · Haobo Yuan, Xiangtai Li, Tao Zhang, Zilong Huang 외

This work presents Sa2VA, the first unified model for dense grounded understanding of both images and videos. Unlike existing multi-modal large language models, which are often limited to specific modalities and tasks, S…

2kLanguage ModelingLanguage ModellingObject+6