paper-with-me

Papers

IC-Portrait: In-Context Matching for View-Consistent Personalized Portrait

2025-01-28 · Han Yang, Enis Simsar, Sotiris Anagnostidis, Yanlong Zang, Thomas Hofmann, Ziwei Liu

Existing diffusion models show great potential for identity-preserving generation. However, personalized portrait generation remains challenging due to the diversity in user profiles, including variations in appearance and lighting conditions. To address these challenges, we propose IC-Portrait, a novel framework designed to accurately encode individual identities for personalized portrait generation. Our key insight is that pre-trained diffusion models are fast learners (e.g.,100 ~ 200 steps) for in-context dense correspondence matching, which motivates the two major designs of our IC-Portrait framework. Specifically, we reformulate portrait generation into two sub-tasks: 1) Lighting-Aware Stitching: we find that masking a high proportion of the input image, e.g., 80%, yields a highly effective self-supervisory representation learning of reference image lighting. 2) View-Consistent Adaptation: we leverage a synthetic view-consistent profile dataset to learn the in-context correspondence. The reference profile can then be warped into arbitrary poses for strong spatial-aligned view conditioning. Coupling these two designs by simply concatenating latents to form ControlNet-like supervision and modeling, enables us to significantly enhance the identity preservation fidelity and stability. Extensive evaluations demonstrate that IC-Portrait consistently outperforms existing state-of-the-art methods both quantitatively and qualitatively, with particularly notable improvements in visual qualities. Furthermore, IC-Portrait even demonstrates 3D-aware relighting capabilities.

📄 PDF Abstract BibTeX arXiv:2501.17159

Code (0)

등록된 구현이 없습니다.

Tasks

Representation Learning

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DiffPortrait360: Consistent Portrait Diffusion for 360 View Synthesis

2025-03-19 · CVPR 2025 1 · Yuming Gu, Phong Tran, Yujian Zheng, Hongyi Xu 외

Generating high-quality 360-degree views of human heads from single-view images is essential for enabling accessible immersive telepresence applications and scalable personalized content creation. While cutting-edge meth…

One-Shot Identity-Preserving Portrait Reenactment

2020-04-26 · Sitao Xiang, Yuming Gu, Pengda Xiang, Mingming He 외

We present a deep learning-based framework for portrait reenactment from a single picture of a target (one-shot) and a video of a driving subject. Existing facial reenactment methods suffer from identity mismatch and pro…

DisentanglementGenerative Adversarial Network

DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis

2023-12-20 · CVPR 2024 1 · Yuming Gu, You Xie, Hongyi Xu, Guoxian Song 외

We present DiffPortrait3D, a conditional diffusion model that is capable of synthesizing 3D-consistent photo-realistic novel views from as few as a single in-the-wild portrait. Specifically, given a single RGB input, we …

Denoising

ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving

2024-04-25 · Jiehui Huang, Xiao Dong, Wenhui Song, Zheng Chong 외

Diffusion-based technologies have made significant strides, particularly in personalized and customized facialgeneration. However, existing methods face challenges in achieving high-fidelity and detailed identity (ID)con…

Diversity

PERSE: Personalized 3D Generative Avatars from A Single Portrait

2024-12-30 · CVPR 2025 1 · Hyunsoo Cha, Inhee Lee, Hanbyul Joo

We present PERSE, a method for building an animatable personalized generative avatar from a reference portrait. Our avatar model enables facial attribute editing in a continuous and disentangled latent space to control e…

Attribute