VOLTA: The Surprising Ineffectiveness of Auxiliary Losses for Calibrated Deep Learning
Uncertainty quantification (UQ) is essential for deploying deep learning models in safety critical applications, yet no consensus exists on which UQ method performs best across different data modalities and distribution shifts. This paper presents a comprehensive benchmark of ten widely used UQ baselines including MC Dropout, SWAG, ensemble methods, temperature scaling, energy based OOD, Mahalanobis, hyperbolic classifiers, ENN, Taylor Sensus, and split conformal prediction against a simplified yet highly effective variant of VOLTA that retains only a deep encoder, learnable prototypes, cross entropy loss, and post hoc temperature scaling. We evaluate all methods on CIFAR 10 (in distribution), CIFAR 100, SVHN, uniform noise (out of distribution), CIFAR 10 C (corruptions), and Tiny ImageNet features (tabular). VOLTA achieves competitive or superior accuracy (up to 0.864 on CIFAR 10), significantly lower expected calibration error (0.010 vs. 0.044 to 0.102 for baselines), and strong OOD detection (AUROC 0.802). Statistical testing over three random seeds shows that VOLTA matches or outperforms most baselines, with ablation studies confirming the importance of adaptive temperature and deep encoders. Our results establish VOLTA as a lightweight, deterministic, and well calibrated alternative to more complex UQ approaches.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Calibrated Value-Aware Model Learning with Stochastic Environment Models
The idea of value-aware model learning, that models should produce accurate value estimates, has gained prominence in model-based reinforcement learning. The MuZero loss, which penalizes a model's value function predicti…
Model-based Reinforcement LearningUnderstanding Representations Pretrained with Auxiliary Losses for Embodied Agent Planning
Pretrained representations from large-scale vision models have boosted the performance of downstream embodied policy learning. We look to understand whether additional self-supervised pretraining on exploration trajector…
Imitation LearningOn the Consistency of Top-k Surrogate Losses
The top-$k$ error is often employed to evaluate performance for challenging classification tasks in computer vision as it is designed to compensate for ambiguity in ground truth labels. This practical success motivates o…
General ClassificationOn the Impact of Different Voltage Unbalance Metrics in Distribution System Optimization
With increasing penetrations of single-phase, rooftop solar PV installations, the relative variations in per-phase loading and associated voltage unbalance are expected to increase. High voltage unbalance may increase ne…
Design of a High Step-up DC-DC Power Converter with Voltage Multiplier Cells and Reduced Losses on Semiconductors for Photovoltaic Systems
A high step up dc dc converter based on an isolated dc dc converter with voltage multiplier cells for photovoltaic systems is essentially introduced in this paper. The proposed converter can provide a high step up voltag…