paper-with-me

홈 › Papers

Appearance Matching Adapter for Exemplar-based Semantic Image Synthesis

2024-12-04 · Siyoon Jin, Jisu Nam, Jiyoung Kim, Dahyun Chung, Yeong-Seok Kim, Joonhyung Park, Heonjeong Chu, Seungryong Kim

Exemplar-based semantic image synthesis aims to generate images aligned with given semantic content while preserving the appearance of an exemplar image. Conventional structure-guidance models, such as ControlNet, are limited in that they cannot directly utilize exemplar images as input, relying instead solely on text prompts to control appearance. Recent tuning-free approaches address this limitation by transferring local appearance from the exemplar image to the synthesized image through implicit cross-image matching in the augmented self-attention mechanism of pre-trained diffusion models. However, these methods face challenges when applied to content-rich scenes with significant geometric deformations, such as driving scenes. In this paper, we propose the Appearance Matching Adapter (AM-Adapter), a learnable framework that enhances cross-image matching within augmented self-attention by incorporating semantic information from segmentation maps. To effectively disentangle generation and matching processes, we adopt a stage-wise training approach. Initially, we train the structure-guidance and generation networks, followed by training the AM-Adapter while keeping the other networks frozen. During inference, we introduce an automated exemplar retrieval method to efficiently select exemplar image-segmentation pairs. Despite utilizing a limited number of learnable parameters, our method achieves state-of-the-art performance, excelling in both semantic alignment preservation and local appearance fidelity. Extensive ablation studies further validate our design choices. Code and pre-trained weights will be publicly available.: https://cvlab-kaist.github.io/AM-Adapter/

📄 PDF Abstract BibTeX arXiv:2412.03150

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationImage SegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

Adapter 설명 없음
ADOPT Please enter a description about the method here
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Delta-Adapter: Scalable Exemplar-Based Image Editing with Single-Pair Supervision

2026-05-08 · Jiacheng Chen, Songze Li, Han Fu, Baoquan Zhao 외 arxiv

Exemplar-based image editing applies a transformation defined by a source-target image pair to a new query image. Existing methods rely on a pair-of-pairs supervision paradigm, requiring two image pairs sharing the same …

Image Editing

A Low-Shot Object Counting Network With Iterative Prototype Adaptation

2022-11-15 · ICCV 2023 1 · Nikola Djukic, Alan Lukezic, Vitjan Zavrtanik, Matej Kristan

We consider low-shot counting of arbitrary semantic categories in the image using only few annotated exemplars (few-shot) or no exemplars (no-shot). The standard few-shot pipeline follows extraction of appearance queries…

Exemplar-Free CountingObjectObject CountingObject Localization

Cross-domain Correspondence Learning for Exemplar-based Image Translation

2020-04-12 · CVPR 2020 6 · Pan Zhang, Bo Zhang, Dong Chen, Lu Yuan 외

We present a general framework for exemplar-based image translation, which synthesizes a photo-realistic image from the input in a distinct domain (e.g., semantic segmentation mask, or edge map, or pose keypoints), given…

Image GenerationImage-to-Image TranslationTranslation

UAE: Universal Anatomical Embedding on Multi-modality Medical Images

2023-11-25 · Xiaoyu Bai, Fan Bai, Xiaofei Huo, Jia Ge 외

Identifying specific anatomical structures (\textit{e.g.}, lesions or landmarks) in medical images plays a fundamental role in medical image analysis. Exemplar-based landmark detection methods are receiving increasing at…

Medical Image AnalysisSelf-Supervised Learning

CoGS: Controllable Generation and Search from Sketch and Style

2022-03-17 · Cusuh Ham, Gemma Canet Tarres, Tu Bui, James Hays 외

We present CoGS, a novel method for the style-conditioned, sketch-driven synthesis of images. CoGS enables exploration of diverse appearance possibilities for a given sketched object, enabling decoupled control over the …

DecoderObject