CrossNet: Latent Cross-Consistency for Unpaired Image Translation
Recent GAN-based architectures have been able to deliver impressive performance on the general task of image-to-image translation. In particular, it was shown that a wide variety of image translation operators may be learned from two image sets, containing images from two different domains, without establishing an explicit pairing between the images. This was made possible by introducing clever regularizers to overcome the under-constrained nature of the unpaired translation problem. In this work, we introduce a novel architecture for unpaired image translation, and explore several new regularizers enabled by it. Specifically, our architecture comprises a pair of GANs, as well as a pair of translators between their respective latent spaces. These cross-translators enable us to impose several regularizing constraints on the learnt image translation operator, collectively referred to as latent cross-consistency. Our results show that our proposed architecture and latent cross-consistency constraints are able to outperform the existing state-of-the-art on a variety of image translation tasks.
Code (0)
등록된 구현이 없습니다.
Tasks
Image-to-Image TranslationTranslationSimilar Papers 제목 키워드 기반
AdaCrossNet: Adaptive Dynamic Loss Weighting for Cross-Modal Contrastive Point Cloud Learning
Manual annotation of large-scale point cloud datasets is laborious due to their irregular structure. While cross-modal contrastive learning methods such as CrossPoint and CrossNet have progressed in utilizing multimodal …
3D Part Segmentation3D Point Cloud Classification3D Point Cloud Linear ClassificationContrastive Learning+1AV-CrossNet: an Audiovisual Complex Spectral Mapping Network for Speech Separation By Leveraging Narrow- and Cross-Band Modeling
Adding visual cues to audio-based speech separation can improve separation performance. This paper introduces AV-CrossNet, an \gls{av} system for speech enhancement, target speaker extraction, and multi-talker speaker se…
Speaker SeparationSpeech EnhancementSpeech SeparationTarget Speaker ExtractionHomomorphic Latent Space Interpolation for Unpaired Image-To-Image Translation
Generative adversarial networks have achieved great success in unpaired image-to-image translation. Cycle consistency allows modeling the relationship between two distinct domains without paired data. In this paper, we p…
Image-to-Image TranslationTranslationXCrossNet: Feature Structure-Oriented Learning for Click-Through Rate Prediction
Click-Through Rate (CTR) prediction is a core task in nowadays commercial recommender systems. Feature crossing, as the mainline of research on CTR prediction, has shown a promising way to enhance predictive performance.…
Click-Through Rate PredictionFeature EngineeringPredictionRecommendation SystemsUnpaired Deep Image Deraining Using Dual Contrastive Learning
Learning single image deraining (SID) networks from an unpaired set of clean and rainy images is practical and valuable as acquiring paired real-world data is almost infeasible. However, without the paired data as the su…
Contrastive LearningImage RestorationRain RemovalSingle Image Deraining