D-LEMA: Deep Learning Ensembles from Multiple Annotations -- Application to Skin Lesion Segmentation
Medical image segmentation annotations suffer from inter- and intra-observer variations even among experts due to intrinsic differences in human annotators and ambiguous boundaries. Leveraging a collection of annotators' opinions for an image is an interesting way of estimating a gold standard. Although training deep models in a supervised setting with a single annotation per image has been extensively studied, generalizing their training to work with datasets containing multiple annotations per image remains a fairly unexplored problem. In this paper, we propose an approach to handle annotators' disagreements when training a deep model. To this end, we propose an ensemble of Bayesian fully convolutional networks (FCNs) for the segmentation task by considering two major factors in the aggregation of multiple ground truth annotations: (1) handling contradictory annotations in the training data originating from inter-annotator disagreements and (2) improving confidence calibration through the fusion of base models' predictions. We demonstrate the superior performance of our approach on the ISIC Archive and explore the generalization performance of our proposed method by cross-dataset evaluation on the PH2 and DermoFit datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Image SegmentationLesion SegmentationMedical Image SegmentationSemantic SegmentationSkin Lesion SegmentationSimilar Papers 제목 키워드 기반
ENHANCE (ENriching Health data by ANnotations of Crowd and Experts): A case study for skin lesion classification
We present ENHANCE, an open dataset with multiple annotations to complement the existing ISIC and PH2 skin lesion classification datasets. This dataset contains annotations of visual ABC (asymmetry, border, colour) featu…
DiagnosticLesion ClassificationMulti-Task LearningSkin Lesion ClassificationMulti-task Ensembles with Crowdsourced Features Improve Skin Lesion Diagnosis
Machine learning has a recognised need for large amounts of annotated data. Due to the high cost of expert annotations, crowdsourcing, where non-experts are asked to label or outline images, has been proposed as an alter…
DiagnosticMulti-Task LearningAn Empirical Study of UMLS Concept Extraction from Clinical Notes using Boolean Combination Ensembles
Our objective in this study is to investigate the behavior of Boolean operators on combining annotation output from multiple Natural Language Processing (NLP) systems across multiple corpora and to assess how filtering b…
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NERStableMask: Refining Causal Masking in Decoder-only Transformer
The decoder-only Transformer architecture with causal masking and relative position encoding (RPE) has become the de facto choice in language modeling. Despite its exceptional performance across various tasks, we have id…
DecoderLanguage ModelingLanguage ModellingPositionEnd-to-end Feature Selection Approach for Learning Skinny Trees
We propose a new optimization-based approach for feature selection in tree ensembles, an important problem in statistics and machine learning. Popular tree ensemble toolkits e.g., Gradient Boosted Trees and Random Forest…
Ensemble LearningFeature CompressionFeature Importancefeature selection+1