paper-with-me

Papers

WeClick: Weakly-Supervised Video Semantic Segmentation with Click Annotations

2021-07-07 · Peidong Liu, Zibin He, Xiyu Yan, Yong Jiang, Shutao Xia, Feng Zheng, Maowei Hu

Compared with tedious per-pixel mask annotating, it is much easier to annotate data by clicks, which costs only several seconds for an image. However, applying clicks to learn video semantic segmentation model has not been explored before. In this work, we propose an effective weakly-supervised video semantic segmentation pipeline with click annotations, called WeClick, for saving laborious annotating effort by segmenting an instance of the semantic class with only a single click. Since detailed semantic information is not captured by clicks, directly training with click labels leads to poor segmentation predictions. To mitigate this problem, we design a novel memory flow knowledge distillation strategy to exploit temporal information (named memory flow) in abundant unlabeled video frames, by distilling the neighboring predictions to the target frame via estimated motion. Moreover, we adopt vanilla knowledge distillation for model compression. In this case, WeClick learns compact video semantic segmentation models with the low-cost click annotations during the training phase yet achieves real-time and accurate models during the inference period. Experimental results on Cityscapes and Camvid show that WeClick outperforms the state-of-the-art methods, increases performance by 10.24% mIoU than baseline, and achieves real-time execution.

📄 PDF Abstract BibTeX arXiv:2107.03088

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge DistillationModel CompressionSegmentationSemantic SegmentationVideo Semantic Segmentation

Methods 이 논문이 사용한 방법론

Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…

Similar Papers 제목 키워드 기반

Weakly Supervised Multiclass Video Segmentation

2014-06-01 · CVPR 2014 6 · Xiao Liu, DaCheng Tao, Mingli Song, Ying Ruan 외

The desire of enabling computers to learn semantic concepts from large quantities of Internet videos has motivated increasing interests on semantic video understanding, while video segmentation is important yet challengi…

SegmentationSemantic SimilaritySemantic Textual SimilarityVideo Segmentation+3

Weakly Supervised Semantic Segmentation using Web-Crawled Videos

2017-01-02 · CVPR 2017 7 · Seunghoon Hong, Donghun Yeo, Suha Kwak, Honglak Lee 외

We propose a novel algorithm for weakly supervised semantic segmentation based on image-level class labels only. In weakly supervised setting, it is commonly observed that trained model overly focuses on discriminative p…

image-classificationImage ClassificationSegmentationSemantic Segmentation+2

Weakly-Supervised Semantic Segmentation using Motion Cues

2016-03-23 · Pavel Tokmakov, Karteek Alahari, Cordelia Schmid

Fully convolutional neural networks (FCNNs) trained on a large number of images with strong pixel-level annotations have become the new state of the art for the semantic segmentation task. While there have been recent at…

Image SegmentationSemantic SegmentationWeakly supervised Semantic SegmentationWeakly-Supervised Semantic Segmentation

Learning to Segment Human by Watching YouTube

2017-10-04 · Xiaodan Liang, Yunchao Wei, Liang Lin, Yunpeng Chen 외

An intuition on human segmentation is that when a human is moving in a video, the video-context (e.g., appearance and motion clues) may potentially infer reasonable mask information for the whole human body. Inspired by …

Human DetectionSegmentationSemantic SegmentationSuperpixels+3

Hierarchical Modeling for Task Recognition and Action Segmentation in Weakly-Labeled Instructional Videos

2021-10-12 · Reza Ghoddoosian, Saif Sayed, Vassilis Athitsos

This paper focuses on task recognition and action segmentation in weakly-labeled instructional videos, where only the ordered sequence of video-level actions is available during training. We propose a two-stream framewor…

Action SegmentationSegmentation