paper-with-me

Papers

DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis

2023-12-20 · CVPR 2024 1 · Yuming Gu, You Xie, Hongyi Xu, Guoxian Song, Yichun Shi, Di Chang, Jing Yang, Linjie Luo

We present DiffPortrait3D, a conditional diffusion model that is capable of synthesizing 3D-consistent photo-realistic novel views from as few as a single in-the-wild portrait. Specifically, given a single RGB input, we aim to synthesize plausible but consistent facial details rendered from novel camera views with retained both identity and facial expression. In lieu of time-consuming optimization and fine-tuning, our zero-shot method generalizes well to arbitrary face portraits with unposed camera views, extreme facial expressions, and diverse artistic depictions. At its core, we leverage the generative prior of 2D diffusion models pre-trained on large-scale image datasets as our rendering backbone, while the denoising is guided with disentangled attentive control of appearance and camera pose. To achieve this, we first inject the appearance context from the reference image into the self-attention layers of the frozen UNets. The rendering view is then manipulated with a novel conditional control module that interprets the camera pose by watching a condition image of a crossed subject from the same view. Furthermore, we insert a trainable cross-view attention module to enhance view consistency, which is further strengthened with a novel 3D-aware noise generation process during inference. We demonstrate state-of-the-art results both qualitatively and quantitatively on our challenging in-the-wild and multi-view benchmarks.

📄 PDF Abstract BibTeX arXiv:2312.13016

Code (1)

FreedomGu/DiffPortrait3D 공식 구현 pytorch

Tasks

Denoising

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DiffPortrait360: Consistent Portrait Diffusion for 360 View Synthesis

2025-03-19 · CVPR 2025 1 · Yuming Gu, Phong Tran, Yujian Zheng, Hongyi Xu 외

Generating high-quality 360-degree views of human heads from single-view images is essential for enabling accessible immersive telepresence applications and scalable personalized content creation. While cutting-edge meth…

Diff-PC: Identity-preserving and 3D-aware Controllable Diffusion for Zero-shot Portrait Customization

2026-01-31 · Yifang Xu, Benxiang Zhai, Chenyu Zhang, Ming Li 외 arxiv

Portrait customization (PC) has recently garnered significant attention due to its potential applications. However, existing PC methods lack precise identity (ID) preservation and face control. To address these tissues, …

HiFi-Portrait: Zero-shot Identity-preserved Portrait Generation with High-fidelity Multi-face Fusion

2025-12-16 · Yifang Xu, Benxiang Zhai, Yunzhuo Sun, Ming Li 외 arxiv

Recent advancements in diffusion-based technologies have made significant strides, particularly in identity-preserved portrait generation (IPG). However, when using multiple reference images from the same ID, existing me…

HiFi-Portrait: Zero-shot Identity-preserved Portrait Generation with High-fidelity Multi-face Fusion

2025-01-01 · CVPR 2025 1 · Yifang Xu, Benxiang Zhai, Yunzhuo Sun, Ming Li 외

Recent advancements in diffusion-based technologies have made significant strides, particularly in identity-preserved portrait generation (IPG). However, when using multiple reference images from the same ID, existin…

High-Fidelity Relightable Monocular Portrait Animation with Lighting-Controllable Video Diffusion Model

2025-02-27 · CVPR 2025 1 · Mingtao Guo, Guanyu Xing, Yanli Liu

Relightable portrait animation aims to animate a static reference portrait to match the head movements and expressions of a driving video while adapting to user-specified or reference lighting conditions. Existing portra…

Portrait Animation