paper-with-me

홈 › Papers

Reducing Overconfident Errors outside the Known Distribution

2019-05-01 · ICLR 2019 5 · Zhizhong Li, Derek Hoiem

Intuitively, unfamiliarity should lead to lack of confidence. In reality, current algorithms often make highly confident yet wrong predictions when faced with unexpected test samples from an unknown distribution different from training. Unlike domain adaptation methods, we cannot gather an "unexpected dataset" prior to test, and unlike novelty detection methods, a best-effort original task prediction is still expected. We compare a number of methods from related fields such as calibration and epistemic uncertainty modeling, as well as two proposed methods that reduce overconfident errors of samples from an unknown novel distribution without drastically increasing evaluation time: (1) G-distillation, training an ensemble of classifiers and then distill into a single model using both labeled and unlabeled examples, or (2) NCR, reducing prediction confidence based on its novelty detection score. Experimentally, we investigate the overconfidence problem and evaluate our solution by creating "familiar" and "novel" test splits, where "familiar" are identically distributed with training and "novel" are not. We discover that calibrating using temperature scaling on familiar data is the best single-model method for improving novel confidence, followed by our proposed methods. In addition, some methods' NLL performance are roughly equivalent to a regularly trained model with certain degree of smoothing. Calibrating can also reduce confident errors, for example, in gender recognition by 95% on demographic groups different from the training data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationNovelty Detection

Similar Papers 제목 키워드 기반

Uncertainty Quantification in Deep Neural Networks through Statistical Inference on Latent Space

2023-05-18 · Luigi Sbailò, Luca M. Ghiringhelli

Uncertainty-quantification methods are applied to estimate the confidence of deep-neural-networks classifiers over their predictions. However, most widely used methods are known to be overconfident. We address this probl…

Uncertainty Quantification

Shaping Parameter Contribution Patterns for Out-of-Distribution Detection

2026-03-07 · Haonan Xu, Yang Yang arxiv

Out-of-distribution (OOD) detection is a well-known challenge due to deep models often producing overconfident. In this paper, we reveal a key insight that trained classifiers tend to rely on sparse parameter contributio…

Out-of-Distribution Detection

Augmentation by Counterfactual Explanation -- Fixing an Overconfident Classifier

2022-10-21 · Sumedha Singla, Nihal Murali, Forough Arabshahi, Sofia Triantafyllou 외

A highly accurate but overconfident model is ill-suited for deployment in critical applications such as healthcare and autonomous driving. The classification outcome should reflect a high uncertainty on ambiguous in-dist…

Autonomous DrivingcounterfactualCounterfactual Explanation

Distribution Calibration for Out-of-Domain Detection with Bayesian Approximation

2022-09-14 · COLING 2022 10 · Yanan Wu, Zhiyuan Zeng, Keqing He, Yutao Mou 외

Out-of-Domain (OOD) detection is a key component in a task-oriented dialog system, which aims to identify whether a query falls outside the predefined supported intent set. Previous softmax-based detection algorithms are…

Out of Distribution (OOD) Detection

OpenMix: Exploring Outlier Samples for Misclassification Detection

2023-03-30 · CVPR 2023 1 · Fei Zhu, Zhen Cheng, Xu-Yao Zhang, Cheng-Lin Liu

Reliable confidence estimation for deep neural classifiers is a challenging yet fundamental requirement in high-stakes applications. Unfortunately, modern deep neural networks are often overconfident for their erroneous …

World Knowledge