Image-to-Image Translation with Disentangled Latent Vectors for Face Editing
We propose an image-to-image translation framework for facial attribute editing with disentangled interpretable latent directions. Facial attribute editing task faces the challenges of targeted attribute editing with controllable strength and disentanglement in the representations of attributes to preserve the other attributes during edits. For this goal, inspired by the latent space factorization works of fixed pretrained GANs, we design the attribute editing by latent space factorization, and for each attribute, we learn a linear direction that is orthogonal to the others. We train these directions with orthogonality constraints and disentanglement losses. To project images to semantically organized latent spaces, we set an encoder-decoder architecture with attention-based skip connections. We extensively compare with previous image translation algorithms and editing with pretrained GAN works. Our extensive experiments show that our method significantly improves over the state-of-the-arts.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeDecoderDisentanglementImage-to-Image TranslationTranslationSimilar Papers 제목 키워드 기반
Smoothing the Disentangled Latent Style Space for Unsupervised Image-to-Image Translation
Image-to-Image (I2I) multi-domain translation models are usually evaluated also using the quality of their semantic interpolation results. However, state-of-the-art models frequently show abrupt changes in the image appe…
Image-to-Image TranslationTranslationUnsupervised Image-To-Image TranslationLearning Disentangled Representations for Image Translation
Recent approaches for unsupervised image translation are strongly reliant on generative adversarial training and architectural locality constraints. Despite their appealing results, it can be easily observed that the lea…
DisentanglementDiversityTranslationDRIT++: Diverse Image-to-Image Translation via Disentangled Representations
Image-to-image translation aims to learn the mapping between two visual domains. There are two main challenges for this task: 1) lack of aligned training pairs and 2) multiple possible outputs from a single input image. …
AttributeDiversityImage-to-Image TranslationPerceptual Distance+1Heredity-aware Child Face Image Generation with Latent Space Disentanglement
Generative adversarial networks have been widely used in image synthesis in recent years and the quality of the generated image has been greatly improved. However, the flexibility to control and decouple facial attribute…
DisentanglementImage GenerationDisentangled Latent Energy-Based Style Translation: An Image-Level Structural MRI Harmonization Framework
Brain magnetic resonance imaging (MRI) has been extensively employed across clinical and research fields, but often exhibits sensitivity to site effects arising from non-biological variations such as differences in field…
Image GenerationTranslation