Learning from Extrinsic and Intrinsic Supervisions for Domain Generalization
The generalization capability of neural networks across domains is crucial for real-world applications. We argue that a generalized object recognition system should well understand the relationships among different images and also the images themselves at the same time. To this end, we present a new domain generalization framework that learns how to generalize across domains simultaneously from extrinsic relationship supervision and intrinsic self-supervision for images from multi-source domains. To be specific, we formulate our framework with feature embedding using a multi-task learning paradigm. Besides conducting the common supervised recognition task, we seamlessly integrate a momentum metric learning task and a self-supervised auxiliary task to collectively utilize the extrinsic supervision and intrinsic supervision. Also, we develop an effective momentum metric learning scheme with K-hard negative mining to boost the network to capture image relationship for domain generalization. We demonstrate the effectiveness of our approach on two standard object recognition benchmarks VLCS and PACS, and show that our methods achieve state-of-the-art performance.
Code (0)
등록된 구현이 없습니다.
Tasks
Anomaly DetectionDomain GeneralizationMetric LearningMulti-Task LearningObject RecognitionSimilar Papers 제목 키워드 기반
Interactive Learning of Intrinsic and Extrinsic Properties for All-day Semantic Segmentation
Scene appearance changes drastically throughout the day. Existing semantic segmentation methods mainly focus on well-lit daytime scenarios and are not well designed to cope with such great appearance changes. Naively usi…
AllAll-day Semantic SegmentationDomain AdaptationSemantic SegmentationHCDG: A Hierarchical Consistency Framework for Domain Generalization on Medical Image Segmentation
Modern deep neural networks struggle to transfer knowledge and generalize across diverse domains when deployed to real-world applications. Currently, domain generalization (DG) is introduced to learn a universal represen…
Data AugmentationDomain GeneralizationImage SegmentationMedical Image Segmentation+3Evaluating Subword Tokenization: Alien Subword Composition and OOV Generalization Challenge
The popular subword tokenizers of current language models, such as Byte-Pair Encoding (BPE), are known not to respect morpheme boundaries, which affects the downstream performance of the models. While many improved token…
text-classificationText ClassificationToward Defining a Domain Complexity Measure Across Domains
Artificial Intelligence (AI) systems planned for deployment in real-world applications frequently are researched and developed in closed simulation environments where all variables are controlled and known to the simulat…
AI AgentJoint Target-Less Intrinsic and Extrinsic Camera-LiDAR Calibration using Deep Point Correspondences
Accurate camera-LiDAR calibration is a prerequisite for robust multi-modal perception in robotics. Recent target-less approaches based on deep point correspondences achieve remarkable performance for extrinsic calibratio…