paper-with-me

홈 › Papers

Can we trust our models? Epistemic calibration in second-order classification

2026-06-09 · Arthur Hoarau arxiv

Uncertainty estimation is critical for deploying machine learning models in high-stakes settings. However, classical calibration only assesses the reliability of predicted probabilities and does not evaluate whether epistemic uncertainty estimates are themselves trustworthy. This limitation is particularly relevant for second-order classification models. We introduce epistemic calibration, a principled criterion that measures whether reported epistemic uncertainty faithfully reflects the dispersion of model predictions around the ground truth. We show that epistemic calibration is a strictly stronger notion than classical calibration and captures failure modes invisible to standard metrics. We relate this work to the existing literature through an impossibility theorem that holds under the epistemic calibration hypothesis. To operationalize this concept, we propose the Expected Epistemic Calibration Error (EECE), which we prove to be a consistent estimator of a True Epistemic Calibration Error (TECE). Experiments across a broad range of uncertainty quantification methods show that epistemic calibration is a coherent and meaningful criterion and reveal substantial differences across methods, despite similar predictive performance.

📄 PDF Abstract BibTeX arXiv:2606.10777

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Minimax Rate of Second-Order Calibration

2026-05-08 · Kamil Ciosek, Banafsheh Rafiee, Sina Ghiassian, Nicolò Felicioni arxiv

We characterize the minimax rate of estimating the second-order calibration error for binary classification, which quantifies whether a higher-order predictor's epistemic-uncertainty estimate matches the conditional vari…

Binary Classification

Is Epistemic Uncertainty Faithfully Represented by Evidential Deep Learning Methods?

2024-02-14 · Mira Jürgens, Nis Meinert, Viktor Bengs, Eyke Hüllermeier 외

Trustworthy ML systems should not only return accurate predictions, but also a reliable representation of their uncertainty. Bayesian methods are commonly used to quantify both aleatoric and epistemic uncertainty, but al…

Deep Learning

Post-Hoc Split-Point Self-Consistency Verification for Efficient, Unified Quantification of Aleatoric and Epistemic Uncertainty in Deep Learning

2025-09-16 · Zhizhong Zhao, Ke Chen arxiv

Uncertainty quantification (UQ) is vital for trustworthy deep learning, yet existing methods are either computationally intensive, such as Bayesian or ensemble methods, or provide only partial, task-specific estimates, s…

Architecting Trust in Artificial Epistemic Agents

2026-03-03 · Nahema Marchal, Stephanie Chan, Matija Franklin, Manon Revel 외 arxiv

Large language models increasingly function as epistemic agents -- entities that can 1) autonomously pursue epistemic goals and 2) actively shape our shared knowledge environment. They curate the information we receive, …

JUCAL: Jointly Calibrating Aleatoric and Epistemic Uncertainty in Classification Tasks

2026-02-23 · Jakob Heiss, Sören Lambrecht, Jakob Weissteiner, Hanna Wutte 외 arxiv

We study post-calibration uncertainty for trained ensembles of classifiers. Specifically, we consider both aleatoric (label noise) and epistemic (model) uncertainty. Among the most popular and widely used calibration met…

Text Classification