High-resolution semantically-consistent image-to-image translation
Deep learning has become one of remote sensing scientists' most efficient computer vision tools in recent years. However, the lack of training labels for the remote sensing datasets means that scientists need to solve the domain adaptation problem to narrow the discrepancy between satellite image datasets. As a result, image segmentation models that are then trained, could better generalize and use an existing set of labels instead of requiring new ones. This work proposes an unsupervised domain adaptation model that preserves semantic consistency and per-pixel quality for the images during the style-transferring phase. This paper's major contribution is proposing the improved architecture of the SemI2I model, which significantly boosts the proposed model's performance and makes it competitive with the state-of-the-art CyCADA model. A second contribution is testing the CyCADA model on the remote sensing multi-band datasets such as WorldView-2 and SPOT-6. The proposed model preserves semantic consistency and per-pixel quality for the images during the style-transferring phase. Thus, the semantic segmentation model, trained on the adapted images, shows substantial performance gain compared to the SemI2I model and reaches similar results as the state-of-the-art CyCADA model. The future development of the proposed method could include ecological domain transfer, {\em a priori} evaluation of dataset quality in terms of data distribution, or exploration of the inner architecture of the domain adaptation model.
Code (0)
등록된 구현이 없습니다.
Tasks
Domain AdaptationImage SegmentationImage-to-Image TranslationSemantic SegmentationTranslationUnsupervised Domain AdaptationVocal Bursts Intensity PredictionSimilar Papers 제목 키워드 기반
Language-agnostic Semantic Consistent Text-to-Image Generation
Recent GAN-based text-to-image generation models have advanced that they can generate photo-realistic images matching semantically with descriptions. However, research on multi-lingual text-to-image generation has not be…
Generative Adversarial NetworkImage GenerationMulti-lingual Text-to-Image GenerationMultilingual Text-to-Image Generation+2SeG-SR: Integrating Semantic Knowledge into Remote Sensing Image Super-Resolution via Vision-Language Model
High-resolution (HR) remote sensing imagery plays a vital role in a wide range of applications, including urban planning and environmental monitoring. However, due to limitations in sensors and data transmission links, t…
Image Super-ResolutionLanguage ModelingLanguage ModellingScene Understanding+1SSCM: A Spatial-Semantic Consistent Model for Multi-Contrast MRI Super-Resolution
Multi-contrast Magnetic Resonance Imaging super-resolution (MC-MRI SR) aims to enhance low-resolution (LR) contrasts leveraging high-resolution (HR) references, shortening acquisition time and improving imaging efficienc…
JAFAR: Jack up Any Feature at Any Resolution
Foundation Vision Encoders have become essential for a wide range of dense vision tasks. However, their low-resolution spatial feature outputs necessitate feature upsampling to produce the high-resolution modalities requ…
Feature UpsamplingSeCo-INR: Semantically Conditioned Implicit Neural Representations for Improved Medical Image Super-Resolution
Implicit Neural Representations (INRs) have recently advanced the field of deep learning due to their ability to learn continuous representations of signals without the need for large training datasets. Although INR meth…
Image Super-ResolutionSemantic SegmentationSuper-Resolution