Style Neophile: Constantly Seeking Novel Styles for Domain Generalization
This paper studies domain generalization via domain-invariant representation learning. Existing methods in this direction suppose that a domain can be characterized by styles of its images, and train a network using style-augmented data so that the network is not biased to particular style distributions. However, these methods are restricted to a finite set of styles since they obtain styles for augmentation from a fixed set of external images or by interpolating those of training data. To address this limitation and maximize the benefit of style augmentation, we propose a new method that synthesizes novel styles constantly during training. Our method manages multiple queues to store styles that have been observed so far, and synthesizes novel styles whose distribution is distinct from the distribution of styles in the queues. The style synthesis process is formulated as a monotone submodular optimization, thus can be conducted efficiently by a greedy algorithm. Extensive experiments on four public benchmarks demonstrate that the proposed method is capable of achieving state-of-the-art domain generalization performance.
Code (0)
등록된 구현이 없습니다.
Tasks
Domain GeneralizationRepresentation LearningSimilar Papers 제목 키워드 기반
Style Aggregated Network for Facial Landmark Detection
Recent advances in facial landmark detection achieve success by learning discriminative features from rich deformation of face shapes and poses. Besides the variance of faces themselves, the intrinsic variance of image s…
Face AlignmentFacial Landmark DetectionPositive Style Accumulation: A Style Screening and Continuous Utilization Framework for Federated DG-ReID
The Federated Domain Generalization for Person re-identification (FedDG-ReID) aims to learn a global server model that can be effectively generalized to source and target domains through distributed source domain data. E…
Person Re-IdentificationDomain GeneralizationStyleAdv: Meta Style Adversarial Training for Cross-Domain Few-Shot Learning
Cross-Domain Few-Shot Learning (CD-FSL) is a recently emerging task that tackles few-shot learning across different domains. It aims at transferring prior knowledge learned on the source dataset to novel target datasets.…
Adversarial AttackCross-Domain Few-Shotcross-domain few-shot learningFew-Shot LearningStyleSinger: Style Transfer for Out-of-Domain Singing Voice Synthesis
Style transfer for out-of-domain (OOD) singing voice synthesis (SVS) focuses on generating high-quality singing voices with unseen styles (such as timbre, emotion, pronunciation, and articulation skills) derived from ref…
QuantizationSinging Voice SynthesisStyle TransferStyle Interleaved Learning for Generalizable Person Re-identification
Domain generalization (DG) for person re-identification (ReID) is a challenging problem, as access to target domain data is not permitted during the training process. Most existing DG ReID methods update the feature extr…
Computational EfficiencyDomain GeneralizationGeneralizable Person Re-identificationMeta-Learning+1