paper-with-me

홈 › Papers

CoCoDiff: Correspondence-Consistent Diffusion Model for Fine-grained Style Transfer

2026-02-16 · Wenbo Nie, Zixiang Li, Renshuai Tao, Bin Wu, Yunchao Wei, Yao Zhao arxiv

Transferring visual style between images while preserving semantic correspondence between similar objects remains a central challenge in computer vision. While existing methods have made great strides, most of them operate at global level but overlook region-wise and even pixel-wise semantic correspondence. To address this, we propose CoCoDiff, a novel training-free and low-cost style transfer framework that leverages pretrained latent diffusion models to achieve fine-grained, semantically consistent stylization. We identify that correspondence cues within generative diffusion models are under-explored and that content consistency across semantically matched regions is often neglected. CoCoDiff introduces a pixel-wise semantic correspondence module that mines intermediate diffusion features to construct a dense alignment map between content and style images. Furthermore, a cycle-consistency module then enforces structural and perceptual alignment across iterations, yielding object and region level stylization that preserves geometry and detail. Despite requiring no additional training or supervision, CoCoDiff delivers state-of-the-art visual quality and strong quantitative results, outperforming methods that rely on extra training or annotations.

📄 PDF Abstract BibTeX arXiv:2602.14464

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic correspondenceStyle Transfer

Similar Papers 제목 키워드 기반

V-Warper: Appearance-Consistent Video Diffusion Personalization via Value Warping

2025-12-13 · Hyunkoo Lee, Wooseok Jang, Jini Yang, Taehwan Kim 외 arxiv

Video personalization aims to generate videos that faithfully reflect a user-provided subject while following a text prompt. However, existing approaches often rely on heavy video-based finetuning or large-scale video da…

MARCO: Navigating the Unseen Space of Semantic Correspondence

2026-04-20 · Claudia Cuttano, Gabriele Trivigno, Carlo Masone, Stefan Roth arxiv

Recent advances in semantic correspondence rely on dual-encoder architectures, combining DINOv2 with diffusion backbones. While accurate, these billion-parameter models generalize poorly beyond training keypoints, reveal…

Semantic correspondence

SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models

2026-05-29 · Olaf Dünkel, Basavaraj Sunagad, Haoran Wang, David T. Hoffmann 외 arxiv

Measuring structured object understanding in vision foundation models remains challenging due to inconsistent evaluation protocols and limited part-level supervision. Semantic correspondence (SC) evaluates this capabilit…

Semantic correspondence3D Pose EstimationImage Matching

EpiDiffVO: Geometry-Aware Epipolar Diffusion for Robust Visual Odometry

2026-05-19 · Prateeth Rao arxiv

Estimating relative pose from image pairs fundamentally requires only a minimal subset of geometrically consistent correspondences. However, most learning-based approaches rely on dense matching or direct regression, lea…

Graph Neural NetworkPose EstimationVisual Odometry

ManiDext: Hand-Object Manipulation Synthesis via Continuous Correspondence Embeddings and Residual-Guided Diffusion

2024-09-14 · Jiajun Zhang, Yuxiang Zhang, Liang An, Mengcheng Li 외

Dynamic and dexterous manipulation of objects presents a complex challenge, requiring the synchronization of hand motions with the trajectories of objects to achieve seamless and physically plausible interactions. In thi…

Denoising