paper-with-me

홈 › Papers

Beyond ImageNet: Understanding Cross-Dataset Robustness of Lightweight Vision Models

2025-11-01 · Weidong Zhang, Pak Lun Kevin Ding, Huan Liu arxiv

Lightweight vision classification models such as MobileNet, ShuffleNet, and EfficientNet are increasingly deployed in mobile and embedded systems, yet their performance has been predominantly benchmarked on ImageNet. This raises critical questions: Do models that excel on ImageNet also generalize across other domains? How can cross-dataset robustness be systematically quantified? And which architectural elements consistently drive generalization under tight resource constraints? Here, we present the first systematic evaluation of 11 lightweight vision models (2.5M parameters), trained under a fixed 100-epoch schedule across 7 diverse datasets. We introduce the Cross-Dataset Score (xScore), a unified metric that quantifies the consistency and robustness of model performance across diverse visual domains. Our results show that (1) ImageNet accuracy does not reliably predict performance on fine-grained or medical datasets, (2) xScore provides a scalable predictor of mobile model performance that can be estimated from just four datasets, and (3) certain architectural components--such as isotropic convolutions with higher spatial resolution and channel-wise attention--promote broader generalization, while Transformer-based blocks yield little additional benefit, despite incurring higher parameter overhead. This study provides a reproducible framework for evaluating lightweight vision models beyond ImageNet, highlights key design principles for mobile-friendly architectures, and guides the development of future models that generalize robustly across diverse application domains.

📄 PDF Abstract BibTeX arXiv:2511.00335

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ConvNets and ImageNet Beyond Accuracy: Understanding Mistakes and Uncovering Biases

2017-11-30 · ECCV 2018 9 · Pierre Stock, Moustapha Cisse

ConvNets and Imagenet have driven the recent success of deep learning for image classification. However, the marked slowdown in performance improvement combined with the lack of robustness of neural networks to adversari…

image-classificationImage Classification

XIMAGENET-12: An Explainable AI Benchmark Dataset for Model Robustness Evaluation

2023-10-12 · Qiang Li, Dan Zhang, Shengzhao Lei, Xun Zhao 외

Despite the promising performance of existing visual models on public benchmarks, the critical assessment of their robustness for real-world applications remains an ongoing challenge. To bridge this gap, we propose an ex…

Classification

On the Robustness of Pretraining and Self-Supervision for a Deep Learning-based Analysis of Diabetic Retinopathy

2021-06-25 · Vignesh Srinivasan, Nils Strodthoff, Jackie Ma, Alexander Binder 외

There is an increasing number of medical use-cases where classification algorithms based on deep neural networks reach performance levels that are competitive with human medical experts. To alleviate the challenges of sm…

Contrastive LearningDiabetic Retinopathy Grading

ImageNet-X: Understanding Model Mistakes with Factor of Variation Annotations

2022-11-03 · Badr Youbi Idrissi, Diane Bouchacourt, Randall Balestriero, Ivan Evtimov 외

Deep learning vision systems are widely deployed across applications where reliability is critical. However, even today's best models can fail to recognize an object when its pose, lighting, or background varies. While e…

Data Augmentation

A Benchmark and Evaluation for Real-World Out-of-Distribution Detection Using Vision-Language Models

2025-01-30 · Shiho Noda, Atsuyuki Miyai, Qing Yu, Go Irie 외

Out-of-distribution (OOD) detection is a task that detects OOD samples during inference to ensure the safety of deployed models. However, conventional benchmarks have reached performance saturation, making it difficult t…

Out-of-Distribution DetectionOut of Distribution (OOD) Detection