DLGAN: Disentangling Label-Specific Fine-Grained Features for Image Manipulation
Recent studies have shown how disentangling images into content and feature spaces can provide controllable image translation/ manipulation. In this paper, we propose a framework to enable utilizing discrete multi-labels to control which features to be disentangled, i.e., disentangling label-specific fine-grained features for image manipulation (dubbed DLGAN). By mapping the discrete label-specific attribute features into a continuous prior distribution, we leverage the advantages of both discrete labels and reference images to achieve image manipulation in a hybrid fashion. For example, given a face image dataset (e.g., CelebA) with multiple discrete fine-grained labels, we can learn to smoothly interpolate a face image between black hair and blond hair through reference images while immediately controlling the gender and age through discrete input labels. To the best of our knowledge, this is the first work that realizes such a hybrid manipulation within a single model. More importantly, it is the first work to achieve image interpolation between two different domains without requiring continuous labels as the supervision. Qualitative and quantitative experiments demonstrate the effectiveness of the proposed method.
Code (1)
Tasks
AttributeImage ManipulationTranslationSimilar Papers 제목 키워드 기반
KD-DLGAN: Data Limited Image Generation via Knowledge Distillation
Generative Adversarial Networks (GANs) rely heavily on large-scale training data for training high-quality image generation models. With limited training data, the GAN discriminator often suffers from severe overfitting …
DiversityImage GenerationKnowledge DistillationFine-Grained Representation Learning via Multi-Level Contrastive Learning without Class Priors
Recent advances in unsupervised representation learning often rely on knowing the number of classes to improve feature extraction and clustering. However, this assumption raises an important question: is the number of cl…
Contrastive LearningDiversityRepresentation Learning3D Guided Fine-Grained Face Manipulation
We present a method for fine-grained face manipulation. Given a face image with an arbitrary expression, our method can synthesize another arbitrary expression by the same person. This is achieved by first fitting a 3D f…
Face ModelCapture Artifacts via Progressive Disentangling and Purifying Blended Identities for Deepfake Detection
The Deepfake technology has raised serious concerns regarding privacy breaches and trust issues. To tackle these challenges, Deepfake detection technology has emerged. Current methods over-rely on the global feature spac…
DeepFake DetectionDisentanglementFace SwappingAttention-based Interactive Disentangling Network for Instance-level Emotional Voice Conversion
Emotional Voice Conversion aims to manipulate a speech according to a given emotion while preserving non-emotion components. Existing approaches cannot well express fine-grained emotional attributes. In this paper, we pr…
Contrastive LearningDisentanglementVoice Conversion