paper-with-me

홈 › Papers

Multi-Rater Calibrated Segmentation Models

2026-05-04 · Meritxell Riera-Marín, Javier García López, Júlia Rodríguez-Comas, Miguel A. González Ballester, Adrian Galdran arxiv

Objective: Accurate probability estimates are essential for the safe deployment of medical image segmentation models in clinical decision-making. However, modern deep segmentation networks are often poorly calibrated, a problem exacerbated when multiple expert annotations exhibit substantial disagreement. While inter-rater variability is typically treated as noise, it provides valuable information about intrinsic annotation ambiguity that must be reflected in model confidence. Methods: We improve the probabilistic calibration of medical image segmentation models by reformulating multi-rater supervision as an ordinal learning problem. Voxel-wise annotator agreement is treated as an ordered target, linking predictive confidence to the empirical variability in training data. This formulation allows the use of ordinal-aware scoring rules, such as the Ranked Probability Score ordinal loss, combined with a standard binary objective to preserve discriminative performance. Results: We evaluated the proposed approach across four public segmentation benchmarks spanning ophthalmology, histopathology, and thoracic imaging. Calibration was assessed using a multi-rater extension of expected calibration error. Results consistently show that ordinal-aware training yields substantially improved calibration with respect to inter-rater agreement without degrading segmentation accuracy. Conclusions: Treating multi-rater annotations as ordered information provides a principled and architecture-agnostic route to more reliable probabilistic segmentation models.

📄 PDF Abstract BibTeX arXiv:2605.02437

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Image Segmentation

Similar Papers 제목 키워드 기반

Learning self-calibrated optic disc and cup segmentation from multi-rater annotations

2022-06-10 · Junde Wu, Huihui Fang, Fangxin Shang, Zhaowei Wang 외

The segmentation of optic disc(OD) and optic cup(OC) from fundus images is an important fundamental task for glaucoma diagnosis. In the clinical practice, it is often necessary to collect opinions from multiple experts t…

Segmentation

Multi-rater Prism: Learning self-calibrated medical image segmentation from multiple raters

2022-12-01 · Junde Wu, Huihui Fang, Yehui Yang, Yuanpei Liu 외

In medical image segmentation, it is often necessary to collect opinions from multiple experts to make the final decision. This clinical routine helps to mitigate individual bias. But when data is multiply annotated, sta…

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation

Learning Calibrated Medical Image Segmentation via Multi-Rater Agreement Modeling

2021-06-19 · CVPR 2021 1 · Wei Ji, Shuang Yu, Junde Wu, Kai Ma 외

In medical image analysis, it is typical to collect multiple annotations, each from a different clinical expert or rater, in the expectation that possible diagnostic errors could be mitigated. Meanwhile, from the com…

DiagnosticImage SegmentationMedical Image AnalysisMedical Image Segmentation+2

TwinTrack: Post-hoc Multi-Rater Calibration for Medical Image Segmentation

2026-04-17 · Tristan Kirscher, Alexandra Ertl, Klaus Maier-Hein, Xavier Coubez 외 arxiv

Pancreatic ductal adenocarcinoma (PDAC) segmentation on contrast-enhanced CT is inherently ambiguous: inter-rater disagreement among experts reflects genuine uncertainty rather than annotation noise. Standard deep learni…

Medical Image Segmentation

Label fusion and training methods for reliable representation of inter-rater uncertainty

2022-02-15 · Andreanne Lemay, Charley Gros, Enamundram Naga Karthik, Julien Cohen-Adad

Medical tasks are prone to inter-rater variability due to multiple factors such as image quality, professional experience and training, or guideline clarity. Training deep learning networks with annotations from multiple…

Segmentation