paper-with-me

홈 › Papers

A framework for benchmarking class-out-of-distribution detection and its application to ImageNet

2023-02-23 · ICLR 2023 2 · Ido Galil, Mohammed Dabbah, Ran El-Yaniv

When deployed for risk-sensitive tasks, deep neural networks must be able to detect instances with labels from outside the distribution for which they were trained. In this paper we present a novel framework to benchmark the ability of image classifiers to detect class-out-of-distribution instances (i.e., instances whose true labels do not appear in the training distribution) at various levels of detection difficulty. We apply this technique to ImageNet, and benchmark 525 pretrained, publicly available, ImageNet-1k classifiers. The code for generating a benchmark for any ImageNet-1k classifier, along with the benchmarks prepared for the above-mentioned 525 models is available at https://github.com/mdabbah/COOD_benchmarking. The usefulness of the proposed framework and its advantage over alternative existing benchmarks is demonstrated by analyzing the results obtained for these models, which reveals numerous novel observations including: (1) knowledge distillation consistently improves class-out-of-distribution (C-OOD) detection performance; (2) a subset of ViTs performs better C-OOD detection than any other model; (3) the language--vision CLIP model achieves good zero-shot detection performance, with its best instance outperforming 96% of all other models evaluated; (4) accuracy and in-distribution ranking are positively correlated to C-OOD detection; and (5) we compare various confidence functions for C-OOD detection. Our companion paper, also published in ICLR 2023 (What Can We Learn From The Selective Prediction And Uncertainty Estimation Performance Of 523 Imagenet Classifiers), examines the uncertainty estimation performance (ranking, calibration, and selective prediction performance) of these classifiers in an in-distribution setting.

📄 PDF Abstract BibTeX arXiv:2302.11893

Code (1)

mdabbah/COOD_benchmarking 공식 구현 pytorch

Tasks

BenchmarkingKnowledge DistillationOut-of-Distribution DetectionOut of Distribution (OOD) Detection

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…

Similar Papers 제목 키워드 기반

Out of Distribution Detection on ImageNet-O

2022-01-23 · Anugya Srivastava, Shriya Jain, Mugdha Thigle

Out of distribution (OOD) detection is a crucial part of making machine learning systems robust. The ImageNet-O dataset is an important tool in testing the robustness of ImageNet trained deep neural networks that are wid…

BenchmarkingOut-of-Distribution DetectionOut of Distribution (OOD) Detection

SemSegBench & DetecBench: Benchmarking Reliability and Generalization Beyond Classification

2025-05-23 · Shashank Agnihotri, David Schader, Jonas Jakubassa, Nico Sharei 외

Reliability and generalization in deep learning are predominantly studied in the context of image classification. Yet, real-world applications in safety-critical domains involve a broader set of semantic tasks, such as s…

BenchmarkingClassificationimage-classificationImage Classification+4

Comparative Benchmarking of Failure Detection Methods in Medical Image Segmentation: Unveiling the Role of Confidence Aggregation

2024-06-05 · Maximilian Zenk, David Zimmerer, Fabian Isensee, Jeremias Traub 외

Semantic segmentation is an essential component of medical image analysis research, with recent deep learning algorithms offering out-of-the-box applicability across diverse datasets. Despite these advancements, segmenta…

BenchmarkingImage SegmentationMedical Image AnalysisMedical Image Segmentation+2

Benchmarking Object Detectors under Real-World Distribution Shifts in Satellite Imagery

2025-03-24 · CVPR 2025 1 · Sara Al-Emadi, Yin Yang, Ferda Ofli

Object detectors have achieved remarkable performance in many applications; however, these deep learning models are typically designed under the i.i.d. assumption, meaning they are trained and evaluated on data sampled f…

BenchmarkingHumanitarianObjectobject-detection+1

ROOD-MRI: Benchmarking the robustness of deep learning segmentation models to out-of-distribution and corrupted data in MRI

2022-03-11 · Lyndon Boone, Mahdi Biparva, Parisa Mojiri Forooshani, Joel Ramirez 외

Deep artificial neural networks (DNNs) have moved to the forefront of medical image analysis due to their success in classification, segmentation, and detection challenges. A principal challenge in large-scale deployment…

BenchmarkingData AugmentationHippocampusImage Segmentation+3