paper-with-me

홈 › Papers

Semantic Data Augmentation based Distance Metric Learning for Domain Generalization

2022-08-02 · Mengzhu Wang, Jianlong Yuan, Qi Qian, Zhibin Wang, Hao Li

Domain generalization (DG) aims to learn a model on one or more different but related source domains that could be generalized into an unseen target domain. Existing DG methods try to prompt the diversity of source domains for the model's generalization ability, while they may have to introduce auxiliary networks or striking computational costs. On the contrary, this work applies the implicit semantic augmentation in feature space to capture the diversity of source domains. Concretely, an additional loss function of distance metric learning (DML) is included to optimize the local geometry of data distribution. Besides, the logits from cross entropy loss with infinite augmentations is adopted as input features for the DML loss in lieu of the deep features. We also provide a theoretical analysis to show that the logits can approximate the distances defined on original features well. Further, we provide an in-depth analysis of the mechanism and rational behind our approach, which gives us a better understanding of why leverage logits in lieu of features can help domain generalization. The proposed DML loss with the implicit augmentation is incorporated into a recent DG method, that is, Fourier Augmented Co-Teacher framework (FACT). Meanwhile, our method also can be easily plugged into various DG methods. Extensive experiments on three benchmarks (Digits-DG, PACS and Office-Home) have demonstrated that the proposed method is able to achieve the state-of-the-art performance.

📄 PDF Abstract BibTeX arXiv:2208.02803

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationDiversityDomain GeneralizationMetric Learning

Similar Papers 제목 키워드 기반

Bridging Domain Gap of Point Cloud Representations via Self-Supervised Geometric Augmentation

2024-09-11 · Li Yu, Hongchao Zhong, Longkun Zou, Ke Chen 외

Recent progress of semantic point clouds analysis is largely driven by synthetic data (e.g., the ModelNet and the ShapeNet), which are typically complete, well-aligned and noisy free. Therefore, representations of those …

Domain AdaptationPoint Cloud ClassificationRepresentation LearningSelf-Supervised Learning+1

Dive into the Resolution Augmentations and Metrics in Low Resolution Face Recognition: A Plain yet Effective New Baseline

2023-02-11 · Xu Ling, Yichen Lu, Wenqi Xu, Weihong Deng 외

Although deep learning has significantly improved Face Recognition (FR), dramatic performance deterioration may occur when processing Low Resolution (LR) faces. To alleviate this, approaches based on unified feature spac…

Face RecognitionGeneral Knowledge

MPA: Multimodal Prototype Augmentation for Few-Shot Learning

2026-02-09 · Liwen Wu, Wei Wang, Lei Zhao, Zhan Gao 외 arxiv

Recently, few-shot learning (FSL) has become a popular task that aims to recognize new classes from only a few labeled examples and has been widely applied in fields such as natural science, remote sensing, and medical i…

Few-Shot Learning

Semantic segmentation of surgical hyperspectral images under geometric domain shifts

2023-03-20 · Jan Sellner, Silvia Seidlitz, Alexander Studier-Fischer, Alessandro Motta 외

Robust semantic segmentation of intraoperative image data could pave the way for automatic surgical scene understanding and autonomous robotic surgery. Geometric domain shifts, however, although common in real-world open…

Organ SegmentationScene SegmentationScene UnderstandingSegmentation+1

Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection

2026-05-13 · Franky George, Muhammad Khalid, Adil Khan arxiv

Night-time pedestrian detection remains challenging because labelled night-time data are limited and large illumination differences make daytime-only trained detectors unreliable. Latent diffusion models (LDMs) provide a…

Image-to-Image TranslationSemantic correspondencePedestrian Detection