Every Annotation Counts: Multi-label Deep Supervision for Medical Image Segmentation
Pixel-wise segmentation is one of the most data and annotation hungry tasks in our field. Providing representative and accurate annotations is often mission-critical especially for challenging medical applications. In this paper, we propose a semi-weakly supervised segmentation algorithm to overcome this barrier. Our approach is based on a new formulation of deep supervision and student-teacher model and allows for easy integration of different supervision signals. In contrast to previous work, we show that care has to be taken how deep supervision is integrated in lower layers and we present multi-label deep supervision as the most important secret ingredient for success. With our novel training regime for segmentation that flexibly makes use of images that are either fully labeled, marked with bounding boxes, just global labels, or not at all, we are able to cut the requirement for expensive labels by 94.22% - narrowing the gap to the best fully supervised baseline to only 5% mean IoU. Our approach is validated by extensive experiments on retinal fluid segmentation and we provide an in-depth analysis of the anticipated effect each annotation type can have in boosting segmentation performance.
Code (1)
Tasks
Image SegmentationMedical Image SegmentationSegmentationSemantic SegmentationWeakly supervised segmentationSimilar Papers 제목 키워드 기반
Fine-Grained Error Analysis and Fair Evaluation of Labeled Spans
The traditional evaluation of labeled spans with precision, recall, and F1-score has undesirable effects due to double penalties. Annotations with incorrect label or boundaries count as two errors instead of one, despite…
ChunkingNERLeveraging Weak Supervision for Cell Localization in Digital Pathology Using Multitask Learning and Consistency Loss
Cell detection and segmentation are integral parts of automated systems in digital pathology. Encoder-decoder networks have emerged as a promising solution for these tasks. However, training of these networks has typical…
Cell DetectionEvery Moment Counts: Dense Detailed Labeling of Actions in Complex Videos
Every moment counts in action recognition. A comprehensive understanding of human activity in video requires labeling every frame according to the actions occurring, placing multiple labels densely over a video sequence.…
Action RecognitionRetrievalTemporal Action LocalizationA flexible model for training action localization with varying levels of supervision
Spatio-temporal action detection in videos is typically addressed in a fully-supervised setup with manual annotation of training videos required at every frame. Since such annotation is extremely tedious and prohibits sc…
Action DetectionAction LocalizationClusteringBootstrapping a 4D LiDAR Annotation Tool from Video Foundation Models
Progress in 4D LiDAR segmentation is bottlenecked by data. Assigning temporally consistent labels across sparse point cloud sequences is costly and hard to scale, and every new task or domain tends to demand fresh dense …
Scene UnderstandingVideo Segmentation