Discriminative Online Learning for Fast Video Object Segmentation
We address the highly challenging problem of video object segmentation. Given only the initial mask, the task is to segment the target in the subsequent frames. In order to effectively handle appearance changes and similar background objects, a robust representation of the target is required. Previous approaches either rely on fine-tuning a segmentation network on the first frame, or employ generative appearance models. Although partially successful, these methods often suffer from impractically low frame rates or unsatisfactory robustness. We propose a novel approach, based on a dedicated target appearance model that is exclusively learned online to discriminate between the target and background image regions. Importantly, we design a specialized loss and customized optimization techniques to enable highly efficient online training. Our light-weight target model is integrated into a carefully designed segmentation network, trained offline to enhance the predictions generated by the target model. Extensive experiments are performed on three datasets. Our approach achieves an overall score of over 70 on YouTube-VOS, while operating at 25 frames per second.
Code (0)
등록된 구현이 없습니다.
Tasks
ObjectOne-shot visual object segmentationSegmentationSemantic SegmentationVideo Object SegmentationVideo Semantic SegmentationSimilar Papers 제목 키워드 기반
D3S -- A Discriminative Single Shot Segmentation Tracker
Template-based discriminative trackers are currently the dominant tracking paradigm due to their robustness, but are restricted to bounding box tracking and a limited range of transformation models, which reduces their l…
ObjectObject TrackingSegmentationSemantic Segmentation+3D3S - A Discriminative Single Shot Segmentation Tracker
Template-based discriminative trackers are currently the dominant tracking paradigm due to their robustness, but are restricted to bounding box tracking and a limited range of transformation models, which reduces their l…
ObjectObject TrackingSegmentationSemantic Segmentation+4A Discriminative Single-Shot Segmentation Network for Visual Object Tracking
Template-based discriminative trackers are currently the dominant tracking paradigm due to their robustness, but are restricted to bounding box tracking and a limited range of transformation models, which reduces their l…
ObjectObject TrackingSegmentationSemantic Segmentation+3Fast and Accurate Online Video Object Segmentation via Tracking Parts
基于视频的目标检测算法研究
Semantic SegmentationSemi-Supervised Video Object SegmentationVideo Object SegmentationVideo Semantic Segmentation+1Joint Inductive and Transductive Learning for Video Object Segmentation
Semi-supervised video object segmentation is a task of segmenting the target object in a video sequence given only a mask annotation in the first frame. The limited information available makes it an extremely challenging…
Inductive LearningObjectSemantic SegmentationSemi-Supervised Video Object Segmentation+3