paper-with-me

홈 › Papers

Two Sides of Miscalibration: Identifying Over and Under-Confidence Prediction for Network Calibration

2023-08-06 · Shuang Ao, Stefan Rueger, Advaith Siddharthan

Proper confidence calibration of deep neural networks is essential for reliable predictions in safety-critical tasks. Miscalibration can lead to model over-confidence and/or under-confidence; i.e., the model's confidence in its prediction can be greater or less than the model's accuracy. Recent studies have highlighted the over-confidence issue by introducing calibration techniques and demonstrated success on various tasks. However, miscalibration through under-confidence has not yet to receive much attention. In this paper, we address the necessity of paying attention to the under-confidence issue. We first introduce a novel metric, a miscalibration score, to identify the overall and class-wise calibration status, including being over or under-confident. Our proposed metric reveals the pitfalls of existing calibration techniques, where they often overly calibrate the model and worsen under-confident predictions. Then we utilize the class-wise miscalibration score as a proxy to design a calibration technique that can tackle both over and under-confidence. We report extensive experiments that show our proposed methods substantially outperforming existing calibration techniques. We also validate our proposed calibration technique on an automatic failure detection task with a risk-coverage curve, reporting that our methods improve failure detection as well as trustworthiness of the model. The code are available at \url{https://github.com/AoShuang92/miscalibration_TS}.

📄 PDF Abstract BibTeX arXiv:2308.03172

Code (1)

aoshuang92/miscalibration_ts 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Pretraining with random noise for uncertainty calibration

2024-12-23 · Jeonghwan Cheon, Se-Bum Paik

Uncertainty calibration is crucial for various machine learning applications, yet it remains challenging. Many models exhibit hallucinations - confident yet inaccurate responses - due to miscalibrated confidence. Here, w…

Discovery of Hidden Miscalibration Regimes

2026-05-13 · Katarzyna Kobalczyk, Mihaela van der Schaar arxiv

Calibration is commonly evaluated by comparing model confidence with its empirical correctness, implicitly treating reliability as a function of the confidence score alone. However, this view can hide substantial structu…

Mind the Confidence Gap: Overconfidence, Calibration, and Distractor Effects in Large Language Models

2025-02-16 · Prateek Chhikara

Large Language Models (LLMs) demonstrate impressive performance across diverse tasks, yet confidence calibration remains a challenge. Miscalibration - where models are overconfident or underconfident - poses risks, parti…

Multiple-choice

Calibrate to Discriminate: Improve In-Context Learning with Label-Free Comparative Inference

2024-10-03 · Wei Cheng, Tianlu Wang, Yanmin Ji, Fan Yang 외

While in-context learning with large language models (LLMs) has shown impressive performance, we have discovered a unique miscalibration behavior where both correct and incorrect predictions are assigned the same level o…

In-Context Learning

Selective Learning: Towards Robust Calibration with Dynamic Regularization

2024-02-13 · Zongbo Han, Yifeng Yang, Changqing Zhang, Linjun Zhang 외

Miscalibration in deep learning refers to there is a discrepancy between the predicted confidence and performance. This problem usually arises due to the overfitting problem, which is characterized by learning everything…