paper-with-me

홈 › Papers

ValUES: A Framework for Systematic Validation of Uncertainty Estimation in Semantic Segmentation

2024-01-16 · Kim-Celine Kahl, Carsten T. Lüth, Maximilian Zenk, Klaus Maier-Hein, Paul F. Jaeger

Uncertainty estimation is an essential and heavily-studied component for the reliable application of semantic segmentation methods. While various studies exist claiming methodological advances on the one hand, and successful application on the other hand, the field is currently hampered by a gap between theory and practice leaving fundamental questions unanswered: Can data-related and model-related uncertainty really be separated in practice? Which components of an uncertainty method are essential for real-world performance? Which uncertainty method works well for which application? In this work, we link this research gap to a lack of systematic and comprehensive evaluation of uncertainty methods. Specifically, we identify three key pitfalls in current literature and present an evaluation framework that bridges the research gap by providing 1) a controlled environment for studying data ambiguities as well as distribution shifts, 2) systematic ablations of relevant method components, and 3) test-beds for the five predominant uncertainty applications: OoD-detection, active learning, failure detection, calibration, and ambiguity modeling. Empirical results on simulated as well as real-world data demonstrate how the proposed framework is able to answer the predominant questions in the field revealing for instance that 1) separation of uncertainty types works on simulated data but does not necessarily translate to real-world data, 2) aggregation of scores is a crucial but currently neglected component of uncertainty methods, 3) While ensembles are performing most robustly across the different downstream tasks and settings, test-time augmentation often constitutes a light-weight alternative. Code is at: https://github.com/IML-DKFZ/values

📄 PDF Abstract BibTeX arXiv:2401.08501

Code (1)

iml-dkfz/values 공식 구현 pytorch

Tasks

Active LearningSemantic Segmentation

Similar Papers 제목 키워드 기반

BeliefPPG: Uncertainty-aware Heart Rate Estimation from PPG signals via Belief Propagation

2023-06-13 · Valentin Bieri, Paul Streli, Berken Utku Demirel, Christian Holz

We present a novel learning-based method that achieves state-of-the-art performance on several heart rate estimation benchmarks extracted from photoplethysmography signals (PPG). We consider the evolution of the heart ra…

Heart rate estimationPhotoplethysmography (PPG) heart rate estimationTime Series AnalysisTime Series Anomaly Detection

Structural System Identification via Validation and Adaptation

2025-06-25 · Cristian López, Keegan J. Moore

Estimating the governing equation parameter values is essential for integrating experimental data with scientific theory to understand, validate, and predict the dynamics of complex systems. In this work, we propose a ne…

parameter estimationUncertainty Quantification

Variational cross-validation of slow dynamical modes in molecular kinetics

2015-03-27

Markov state models (MSMs) are a widely used method for approximating the eigenspectrum of the molecular dynamics propagator, yielding insight into the long-timescale statistical kinetics and slow dynamical modes of biom…

Can We Validate Counterfactual Estimations in the Presence of General Network Interference?

2025-02-03 · Sadegh Shirani, Yuwei Luo, William Overman, Ruoxuan Xiong 외

In experimental settings with network interference, a unit's treatment can influence outcomes of other units, challenging both causal effect estimation and its validation. Classic validation approaches fail as outcomes a…

Causal Inferencecounterfactual

Validation-Induced Shapley Shifts: How Validation Structure Distorts Data Valuation

2026-07-04 · Yinan Shen, Ziao Yang, Hongfu Liu arxiv

Shapley values are widely used to attribute value to training data based on their marginal contribution to performance on a validation set. Existing practice often assumes these values are stable once the training data a…