paper-with-me

홈 › Papers

DiffusionRig: Learning Personalized Priors for Facial Appearance Editing

2023-04-13 · CVPR 2023 1 · Zheng Ding, Xuaner Zhang, Zhihao Xia, Lars Jebe, Zhuowen Tu, Xiuming Zhang

We address the problem of learning person-specific facial priors from a small number (e.g., 20) of portrait photos of the same person. This enables us to edit this specific person's facial appearance, such as expression and lighting, while preserving their identity and high-frequency facial details. Key to our approach, which we dub DiffusionRig, is a diffusion model conditioned on, or "rigged by," crude 3D face models estimated from single in-the-wild images by an off-the-shelf estimator. On a high level, DiffusionRig learns to map simplistic renderings of 3D face models to realistic photos of a given person. Specifically, DiffusionRig is trained in two stages: It first learns generic facial priors from a large-scale face dataset and then person-specific priors from a small portrait photo collection of the person of interest. By learning the CGI-to-photo mapping with such personalized priors, DiffusionRig can "rig" the lighting, facial expression, head pose, etc. of a portrait photo, conditioned only on coarse 3D models while preserving this person's identity and other high-frequency characteristics. Qualitative and quantitative experiments show that DiffusionRig outperforms existing approaches in both identity preservation and photorealism. Please see the project website: https://diffusionrig.github.io for the supplemental material, video, code, and data.

📄 PDF Abstract BibTeX arXiv:2304.06711

Code (1)

adobe-research/diffusion-rig 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

StrokeFaceNeRF: Stroke-based Facial Appearance Editing in Neural Radiance Field

2024-01-01 · CVPR 2024 1 · Xiao-Juan Li, Dingxi Zhang, Shu-Yu Chen, Feng-Lin Liu

Current 3D-aware facial NeRF generation approaches control the facial appearance by text lighting conditions or reference images limiting precise manipulation of local facial regions and interactivity. Color stroke a…

3D geometryNeRF

NeRFFaceEditing: Disentangled Face Editing in Neural Radiance Fields

2022-11-15 · Kaiwen Jiang, Shu-Yu Chen, Feng-Lin Liu, Hongbo Fu 외

Recent methods for synthesizing 3D-aware face images have achieved rapid development thanks to neural radiance fields, allowing for high quality and fast inference speed. However, existing solutions for editing facial ge…

DecoderDisentanglement

Beyond Facial Consistency: Personalized Person Image Generation with Holistic Identity Preservation

2026-07-28 · Yuxuan Xiao, Shanshan Zhang, Jian Yang, Shengcai Liao arxiv

Personalized person image generation requires preserving subject identity across both local facial details and broader appearance cues. Existing methods typically emphasize only one level of identity information, leading…

Image Generation

Monocular and Generalizable Gaussian Talking Head Animation

2025-04-01 · CVPR 2025 1 · Shengjie Gong, Haojie Li, Jiapeng Tang, Dongming Hu 외

In this work, we introduce Monocular and Generalizable Gaussian Talking Head Animation (MGGTalk), which requires monocular datasets and generalizes to unseen identities without personalized re-training. Compared with pre…

3DGSDepth Estimation

Exemplar-based Generative Facial Editing

2020-05-31 · Jingtao Guo, Yi Liu, Zhenzhen Qian, Zuowei Zhou

Image synthesis has witnessed substantial progress due to the increasing power of generative model. This paper we propose a novel generative approach for exemplar based facial editing in the form of the region inpainting…

AttributeFacial EditingImage Generation