paper-with-me

홈 › Papers

Scaling of Class-wise Training Losses for Post-hoc Calibration

2023-06-19 · Seungjin Jung, Seungmo Seo, Yonghyun Jeong, Jongwon Choi

The class-wise training losses often diverge as a result of the various levels of intra-class and inter-class appearance variation, and we find that the diverging class-wise training losses cause the uncalibrated prediction with its reliability. To resolve the issue, we propose a new calibration method to synchronize the class-wise training losses. We design a new training loss to alleviate the variance of class-wise training losses by using multiple class-wise scaling factors. Since our framework can compensate the training losses of overfitted classes with those of under-fitted classes, the integrated training loss is preserved, preventing the performance drop even after the model calibration. Furthermore, our method can be easily employed in the post-hoc calibration methods, allowing us to use the pre-trained model as an initial model and reduce the additional computation for model calibration. We validate the proposed framework by employing it in the various post-hoc calibration methods, which generally improves calibration performance while preserving accuracy, and discover through the investigation that our approach performs well with unbalanced datasets and untuned hyperparameters.

📄 PDF Abstract BibTeX arXiv:2306.10989

Code (1)

seungjinjung/sctl 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Soft Calibration Objectives for Neural Networks

2021-07-30 · NeurIPS 2021 12 · Archit Karandikar, Nicholas Cain, Dustin Tran, Balaji Lakshminarayanan 외

Optimal decision making requires that classifiers produce uncertainty estimates consistent with their empirical accuracy. However, deep neural networks are often under- or over-confident in their predictions. Consequentl…

Decision Making

Improving Calibration by Relating Focal Loss, Temperature Scaling, and Properness

2024-08-21 · Viacheslav Komisarenko, Meelis Kull

Proper losses such as cross-entropy incentivize classifiers to produce class probabilities that are well-calibrated on the training data. Due to the generalization gap, these classifiers tend to become overconfident on t…

image-classificationImage Classification

RankT5: Fine-Tuning T5 for Text Ranking with Ranking Losses

2022-10-12 · Honglei Zhuang, Zhen Qin, Rolf Jagerman, Kai Hui 외

Recently, substantial progress has been made in text ranking based on pretrained language models such as BERT. However, there are limited studies on how to leverage more powerful sequence-to-sequence models such as T5. E…

Decoder

ChopGrad: Pixel-Wise Losses for Latent Video Diffusion via Truncated Backpropagation

2026-03-18 · Dmitriy Rivkin, Parker Ewen, Lili Gao, Julian Ost 외 arxiv

Recent video diffusion models achieve high-quality generation through recurrent frame processing where each frame generation depends on previous frames. However, this recurrent mechanism means that training such models i…

Video Super-ResolutionVideo EnhancementVideo GenerationVideo Inpainting

PTQ-SL: Exploring the Sub-layerwise Post-training Quantization

2021-10-15 · Zhihang Yuan, Yiqi Chen, Chenhao Xue, Chenguang Zhang 외

Network quantization is a powerful technique to compress convolutional neural networks. The quantization granularity determines how to share the scaling factors in weights, which affects the performance of network quanti…

Quantization