paper-with-me

홈 › Papers

Dual-Domain Image Synthesis using Segmentation-Guided GAN

2022-04-19 · Dena Bazazian, Andrew Calway, Dima Damen

We introduce a segmentation-guided approach to synthesise images that integrate features from two distinct domains. Images synthesised by our dual-domain model belong to one domain within the semantic mask, and to another in the rest of the image - smoothly integrated. We build on the successes of few-shot StyleGAN and single-shot semantic segmentation to minimise the amount of training required in utilising two domains. The method combines a few-shot cross-domain StyleGAN with a latent optimiser to achieve images containing features of two distinct domains. We use a segmentation-guided perceptual loss, which compares both pixel-level and activations between domain-specific and dual-domain synthetic images. Results demonstrate qualitatively and quantitatively that our model is capable of synthesising dual-domain images on a variety of objects (faces, horses, cats, cars), domains (natural, caricature, sketches) and part-based masks (eyes, nose, mouth, hair, car bonnet). The code is publicly available at: https://github.com/denabazazian/Dual-Domain-Synthesis.

📄 PDF Abstract BibTeX arXiv:2204.09015

Code (1)

denabazazian/Dual-Domain-Synthesis 공식 구현 pytorch

Tasks

CaricatureImage GenerationSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
StyleGAN 설명 없음
Adaptive Instance Normalization 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
Feedforward Network A Feedforward Network, or a Multilayer Perceptron (MLP), is a neural network with solely densely connected layers. This is the classic neural network architecture of the…

Similar Papers 제목 키워드 기반

Retrieval-guided Cross-view Image Synthesis

2024-11-29 · Hongji Yang, Yiru Li, Yingying Zhu

Information retrieval techniques have demonstrated exceptional capabilities in identifying semantic similarities across diverse domains through robust feature representations. However, their potential in guiding synthesi…

Contrastive LearningDiversityImage GenerationInformation Retrieval+3

Less is More: Unsupervised Mask-guided Annotated CT Image Synthesis with Minimum Manual Segmentations

2023-03-19 · Xiaodan Xing, Giorgos Papanastasiou, Simon Walsh, Guang Yang

As a pragmatic data augmentation tool, data synthesis has generally returned dividends in performance for deep learning based medical image analysis. However, generating corresponding segmentation masks for synthetic med…

Data AugmentationImage GenerationMedical Image AnalysisSegmentation

Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation

2025-11-05 · Pengyu Jie, Wanquan Liu, Rui He, Yihui Wen 외 arxiv

Augmentation for dense prediction typically relies on either sample mixing or generative synthesis. Mixing improves robustness but misaligned masks yield soft label ambiguity. Diffusion synthesis increases apparent diver…

Image Segmentation

CrackSegFlow: Controllable Flow Matching Synthesis for Generalizable Crack Segmentation with a 50K Image-Mask Benchmark

2026-01-07 · Babak Asadi, Peiyang Wu, Mani Golparvar-Fard, Ramez Hajj arxiv

Defect segmentation is central to computer vision based inspection of infrastructure assets during both construction and operation. However, deployment remains limited due to scarce pixel-level labels and domain shift ac…

Crack Segmentation

Towards Fair and Robust Face Parsing for Generative AI: A Multi-Objective Approach

2025-02-06 · Sophia J. Abraham, Jonathan D. Hauenstein, Walter J. Scheirer

Face parsing is a fundamental task in computer vision, enabling applications such as identity verification, facial editing, and controllable image synthesis. However, existing face parsing models often lack fairness and …

Face GenerationFace ParsingFacial EditingFairness+2