Learning to Better Segment Objects from Unseen Classes with Unlabeled Videos
The ability to localize and segment objects from unseen classes would open the door to new applications, such as autonomous object learning in active vision. Nonetheless, improving the performance on unseen classes requires additional training data, while manually annotating the objects of the unseen classes can be labor-extensive and expensive. In this paper, we explore the use of unlabeled video sequences to automatically generate training data for objects of unseen classes. It is in principle possible to apply existing video segmentation methods to unlabeled videos and automatically obtain object masks, which can then be used as a training set even for classes with no manual labels available. However, our experiments show that these methods do not perform well enough for this purpose. We therefore introduce a Bayesian method that is specifically designed to automatically create such a training set: Our method starts from a set of object proposals and relies on (non-realistic) analysis-by-synthesis to select the correct ones by performing an efficient optimization over all the frames simultaneously. Through extensive experiments, we show that our method can generate a high-quality training set which significantly boosts the performance of segmenting objects of unseen classes. We thus believe that our method could open the door for open-world instance segmentation using abundant Internet videos.
Code (0)
등록된 구현이 없습니다.
Tasks
Instance SegmentationObjectOpen-World Instance SegmentationSemantic SegmentationVideo SegmentationVideo Semantic SegmentationSimilar Papers 제목 키워드 기반
Segmenting Known Objects and Unseen Unknowns without Prior Knowledge
Panoptic segmentation methods assign a known class to each pixel given in input. Even for state-of-the-art approaches, this inevitably enforces decisions that systematically lead to wrong predictions for objects outside …
Instance SegmentationObject DetectionPanoptic SegmentationScene Understanding+1Re-Evaluating the Impact of Unseen-Class Unlabeled Data on Semi-Supervised Learning Model
Semi-supervised learning (SSL) effectively leverages unlabeled data and has been proven successful across various fields. Current safe SSL methods believe that unseen classes in unlabeled data harm the performance of SSL…
A New Few-shot Segmentation Network Based on Class Representation
This paper studies few-shot segmentation, which is a task of predicting foreground mask of unseen classes by a few of annotations only, aided by a set of rich annotations already existed. The existing methods mainly focu…
SegmentationPrototypical Matching and Open Set Rejection for Zero-Shot Semantic Segmentation
The deep learning methods in addressing semantic segmentation typically demand vast amount of pixel-wise annotated training samples. In this work, we present zero-shot semantic segmentation, which aims to identify no…
SegmentationSemantic SegmentationZero-Shot Semantic SegmentationDoUnseen: Tuning-Free Class-Adaptive Object Detection of Unseen Objects for Robotic Grasping
How can we segment varying numbers of objects where each specific object represents its own separate class? To make the problem even more realistic, how can we add and delete classes on the fly without retraining or fine…
Objectobject-detectionObject DetectionRobotic Grasping+4