DUNIT: Detection-Based Unsupervised Image-to-Image Translation
Image-to-image translation has made great strides in recent years, with current techniques being able to handle unpaired training images and to account for the multi-modality of the translation problem. Despite this, most methods treat the image as a whole, which makes the results they produce for content-rich scenes less realistic. In this paper, we introduce a Detection-based Unsupervised Image-to-image Translation (DUNIT) approach that explicitly accounts for the object instances in the translation process. To this end, we extract separate representations for the global image and for the instances, which we then fuse into a common representation from which we generate the translated image. This allows us to preserve the detailed content of object instances, while still modeling the fact that we aim to produce an image of a single consistent scene. We introduce an instance consistency loss to maintain the coherence between the detections. Furthermore, by incorporating a detector into our architecture, we can still exploit object instances at test time. As evidenced by our experiments, this allows us to outperform the state-of-the-art unsupervised image-to-image translation methods. Furthermore, our approach can also be used as an unsupervised domain adaptation strategy for object detection, and it also achieves state-of-the-art performance on this task.
Code (1)
Tasks
Domain AdaptationImage-to-Image TranslationObjectobject-detectionObject DetectionTranslationUnsupervised Domain AdaptationUnsupervised Image-To-Image TranslationSimilar Papers 제목 키워드 기반
Few-Shot Unsupervised Image-to-Image Translation on complex scenes
Unsupervised image-to-image translation methods have received a lot of attention in the last few years. Multiple techniques emerged tackling the initial challenge from different perspectives. Some focus on learning as mu…
Image-to-Image TranslationObjectobject-detectionObject Detection+2UCDFormer: Unsupervised Change Detection Using a Transformer-driven Image Translation
Change detection (CD) by comparing two bi-temporal images is a crucial task in remote sensing. With the advantages of requiring no cumbersome labeled change information, unsupervised CD has attracted extensive attention …
Change DetectionTranslationShow, Attend and Translate: Unsupervised Image Translation with Self-Regularization and Attention
Image translation between two domains is a class of problems aiming to learn mapping from an input image in the source domain to an output image in the target domain. It has been applied to numerous domains, such as data…
Data AugmentationDomain AdaptationSaliency DetectionTranslationBrainomaly: Unsupervised Neurologic Disease Detection Utilizing Unannotated T1-weighted Brain MR Images
Harnessing the power of deep neural networks in the medical imaging domain is challenging due to the difficulties in acquiring large annotated datasets, especially for rare diseases, which involve high costs, time, and e…
Alzheimer's Disease DetectionAnomaly DetectionImage-to-Image TranslationModel Selection+1Semantically Robust Unsupervised Image Translation for Paired Remote Sensing Images
Image translation for change detection or classification in bi-temporal remote sensing images is unique. Although it can acquire paired images, it is still unsupervised. Moreover, strict semantic preservation in translat…
Change DetectionImage-to-Image TranslationTranslationUnsupervised Image-To-Image Translation