paper-with-me

홈 › Papers

On-Device Domain Generalization

2022-09-15 · Kaiyang Zhou, Yuanhan Zhang, Yuhang Zang, Jingkang Yang, Chen Change Loy, Ziwei Liu

We present a systematic study of domain generalization (DG) for tiny neural networks. This problem is critical to on-device machine learning applications but has been overlooked in the literature where research has been merely focused on large models. Tiny neural networks have much fewer parameters and lower complexity and therefore should not be trained the same way as their large counterparts for DG applications. By conducting extensive experiments, we find that knowledge distillation (KD), a well-known technique for model compression, is much better for tackling the on-device DG problem than conventional DG methods. Another interesting observation is that the teacher-student gap on out-of-distribution data is bigger than that on in-distribution data, which highlights the capacity mismatch issue as well as the shortcoming of KD. We further propose a method called out-of-distribution knowledge distillation (OKD) where the idea is to teach the student how the teacher handles out-of-distribution data synthesized via disruptive data augmentation. Without adding any extra parameter to the model -- hence keeping the deployment cost unchanged -- OKD significantly improves DG performance for tiny neural networks in a variety of on-device DG scenarios for image and speech applications. We also contribute a scalable approach for synthesizing visual domain shifts, along with a new suite of DG datasets to complement existing testbeds.

📄 PDF Abstract BibTeX arXiv:2209.07521

Code (2)

kaiyangzhou/on-device-dg 공식 구현 pytorch
KaiyangZhou/mixstyle-release pytorch

Tasks

Data AugmentationDomain GeneralizationKnowledge DistillationModel Compression

Methods 이 논문이 사용한 방법론

Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…

Similar Papers 제목 키워드 기반

Discussion on domain generalization in the cross-device speaker verification system

2021-10-01 · ROCLING 2021 10 · Wei-Ting Lin, Yu-jia Zhang, Chia-Ping Chen, Chung-Li Lu 외

In this paper, we use domain generalization to improve the performance of the cross-device speaker verification system. Based on a trainable speaker verification system, we use domain generalization algorithms to fine-tu…

Domain GeneralizationSpeaker Verification

Domain Information Control at Inference Time for Acoustic Scene Classification

2023-06-13 · Shahed Masoudian, Khaled Koutini, Markus Schedl, Gerhard Widmer 외

Domain shift is considered a challenge in machine learning as it causes significant degradation of model performance. In the Acoustic Scene Classification task (ASC), domain shift is mainly caused by different recording …

Acoustic Scene ClassificationDomain GeneralizationScene Classification

InvNorm: Domain Generalization for Object Detection in Gastrointestinal Endoscopy

2022-05-05 · Weichen Fan, Yuanbo Yang, Kunpeng Qiu, Shuo Wang 외

Domain Generalization is a challenging topic in computer vision, especially in Gastrointestinal Endoscopy image analysis. Due to several device limitations and ethical reasons, current open-source datasets are typically …

Domain GeneralizationEthicsObjectobject-detection+1

Domain Generalization with Relaxed Instance Frequency-wise Normalization for Multi-device Acoustic Scene Classification

2022-06-24 · Byeonggeun Kim, Seunghan Yang, Jangho Kim, Hyunsin Park 외

While using two-dimensional convolutional neural networks (2D-CNNs) in image processing, it is possible to manipulate domain information using channel statistics, and instance normalization has been a promising way to ge…

Acoustic Scene ClassificationDomain GeneralizationScene Classification

Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions

2026-05-28 · Heejoon Koo, Yoon Tae Kim, Miika Toikkanen, June-Woo Kim arxiv

AI-driven respiratory sound classification (RSC) is promising for automated pulmonary disease detection, yet multi-site deployment is hindered by inter-stethoscope variability. We introduce a federated domain generalizat…

Domain GeneralizationFederated LearningData Augmentation