paper-with-me

홈 › Papers

Uncertainty-Aware ControlNet: Bridging Domain Gaps with Synthetic Image Generation

2025-10-13 · Joshua Niemeijer, Jan Ehrhardt, Heinz Handels, Hristina Uzunova arxiv

Generative Models are a valuable tool for the controlled creation of high-quality image data. Controlled diffusion models like the ControlNet have allowed the creation of labeled distributions. Such synthetic datasets can augment the original training distribution when discriminative models, like semantic segmentation, are trained. However, this augmentation effect is limited since ControlNets tend to reproduce the original training distribution. This work introduces a method to utilize data from unlabeled domains to train ControlNets by introducing the concept of uncertainty into the control mechanism. The uncertainty indicates that a given image was not part of the training distribution of a downstream task, e.g., segmentation. Thus, two types of control are engaged in the final network: an uncertainty control from an unlabeled dataset and a semantic control from the labeled dataset. The resulting ControlNet allows us to create annotated data with high uncertainty from the target domain, i.e., synthetic data from the unlabeled distribution with labels. In our scenario, we consider retinal OCTs, where typically high-quality Spectralis images are available with given ground truth segmentations, enabling the training of segmentation networks. The recent development in Home-OCT devices, however, yields retinal OCTs with lower quality and a large domain shift, such that out-of-the-pocket segmentation networks cannot be applied for this type of data. Synthesizing annotated images from the Home-OCT domain using the proposed approach closes this gap and leads to significantly improved segmentation results without adding any further supervision. The advantage of uncertainty-guidance becomes obvious when compared to style transfer: it enables arbitrary domain shifts without any strict learning of an image style. This is also demonstrated in a traffic scene experiment.

📄 PDF Abstract BibTeX arXiv:2510.11346

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic SegmentationImage GenerationStyle Transfer

Similar Papers 제목 키워드 기반

Bridging Synthetic-to-Real Gaps: Frequency-Aware Perturbation and Selection for Single-shot Multi-Parametric Mapping Reconstruction

2025-03-05 · Linyu Fan, Che Wang, Ming Ye, Qizhi Yang 외

Data-centric artificial intelligence (AI) has remarkably advanced medical imaging, with emerging methods using synthetic data to address data scarcity while introducing synthetic-to-real gaps. Unsupervised domain adaptat…

Domain AdaptationUnsupervised Domain Adaptation

Bridging the Vision-Brain Gap with an Uncertainty-Aware Blur Prior

2025-03-06 · CVPR 2025 1 · Haitao Wu, Qing Li, Changqing Zhang, Zhen He 외

Can our brain signals faithfully reflect the original visual stimuli, even including high-frequency details? Although human perceptual and cognitive capacities enable us to process and remember visual information, these …

Image Retrieval

SPAC-Net: Synthetic Pose-aware Animal ControlNet for Enhanced Pose Estimation

2023-05-29 · Le Jiang, Sarah Ostadabbas

Animal pose estimation has become a crucial area of research, but the scarcity of annotated data is a significant challenge in developing accurate models. Synthetic data has emerged as a promising alternative, but it fre…

Animal Pose EstimationEdge DetectionPose EstimationStyle Transfer

S2R-ViT for Multi-Agent Cooperative Perception: Bridging the Gap from Simulation to Reality

2023-07-16 · Jinlong Li, Runsheng Xu, Xinyu Liu, Baolu Li 외

Due to the lack of enough real multi-agent data and time-consuming of labeling, existing multi-agent cooperative perception algorithms usually select the simulated sensor data for training and validating. However, the pe…

3D Object Detectionobject-detectionObject DetectionTransfer Learning

PromptSync: Bridging Domain Gaps in Vision-Language Models through Class-Aware Prototype Alignment and Discrimination

2024-04-11 · Anant Khandelwal

The potential for zero-shot generalization in vision-language (V-L) models such as CLIP has spurred their widespread adoption in addressing numerous downstream tasks. Previous methods have employed test-time prompt tunin…

Contrastive LearningDomain GeneralizationZero-shot Generalization