Learning Disentangled Representations for Image Translation
Recent approaches for unsupervised image translation are strongly reliant on generative adversarial training and architectural locality constraints. Despite their appealing results, it can be easily observed that the learned class and content representations are entangled which often hurts the translation performance. To this end, we propose OverLORD, for learning disentangled representations for the image class and attributes, utilizing latent optimization and carefully designed content and style bottlenecks. We further argue that the commonly used adversarial optimization can be decoupled from representation disentanglement and be applied at a later stage of the training to increase the perceptual quality of the generated images. Based on these principles, our model learns significantly more disentangled representations and achieves higher translation quality and greater output diversity than state-of-the-art methods.
Code (0)
등록된 구현이 없습니다.
Tasks
DisentanglementDiversityTranslationSimilar Papers 제목 키워드 기반
Rotation and Translation Invariant Representation Learning with Implicit Neural Representations
In many computer vision applications, images are acquired with arbitrary or random rotations and translations, and in such setups, it is desirable to obtain semantic representations disentangled from the image orientatio…
ClusteringRepresentation LearningTranslationImage Generation and Translation with Disentangled Representations
Generative models have made significant progress in the tasks of modeling complex data distributions such as natural images. The introduction of Generative Adversarial Networks (GANs) and auto-encoders lead to the possib…
Conditional Image GenerationFace GenerationImage GenerationImage-to-Image Translation+2Self-Supervised 2D Image to 3D Shape Translation with Disentangled Representations
We present a framework to translate between 2D image views and 3D object shapes. Recent progress in deep learning enabled us to learn structure-aware representations from a scene. However, the existing literature assumes…
Image to 3DTranslationDRIT++: Diverse Image-to-Image Translation via Disentangled Representations
Image-to-image translation aims to learn the mapping between two visual domains. There are two main challenges for this task: 1) lack of aligned training pairs and 2) multiple possible outputs from a single input image. …
AttributeDiversityImage-to-Image TranslationPerceptual Distance+1Learning Disentangled Representations of Satellite Image Time Series
In this paper, we investigate how to learn a suitable representation of satellite image time series in an unsupervised manner by leveraging large amounts of unlabeled data. Additionally , we aim to disentangle the repres…
Change Detectionimage-classificationImage ClassificationImage Retrieval+7