paper-with-me

Papers

FaceDNeRF: Semantics-Driven Face Reconstruction, Prompt Editing and Relighting with Diffusion Models

2023-06-01 · NeurIPS 2023 11 · Hao Zhang, Yanbo Xu, Tianyuan Dai, Yu-Wing Tai, Chi-Keung Tang

The ability to create high-quality 3D faces from a single image has become increasingly important with wide applications in video conferencing, AR/VR, and advanced video editing in movie industries. In this paper, we propose Face Diffusion NeRF (FaceDNeRF), a new generative method to reconstruct high-quality Face NeRFs from single images, complete with semantic editing and relighting capabilities. FaceDNeRF utilizes high-resolution 3D GAN inversion and expertly trained 2D latent-diffusion model, allowing users to manipulate and construct Face NeRFs in zero-shot learning without the need for explicit 3D data. With carefully designed illumination and identity preserving loss, as well as multi-modal pre-training, FaceDNeRF offers users unparalleled control over the editing process enabling them to create and edit face NeRFs using just single-view images, text prompts, and explicit target lighting. The advanced features of FaceDNeRF have been designed to produce more impressive results than existing 2D editing approaches that rely on 2D segmentation maps for editable attributes. Experiments show that our FaceDNeRF achieves exceptionally realistic results and unprecedented flexibility in editing compared with state-of-the-art 3D face reconstruction and editing methods. Our code will be available at https://github.com/BillyXYB/FaceDNeRF.

📄 PDF Abstract BibTeX arXiv:2306.00783

Code (2)

billyxyb/facednerf 공식 구현 pytorch
billyxyb/fdnerf 공식 구현 pytorch

Tasks

3D Face ReconstructionFace ReconstructionNeRFVideo EditingZero-Shot Learning

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

CLIP-FTI: Fine-Grained Face Template Inversion via CLIP-Driven Attribute Conditioning

2025-12-17 · Longchen Dai, Zixuan Shen, Zhiheng Zhou, Peipeng Yu 외 arxiv

Face recognition systems store face templates for efficient matching. Once leaked, these templates pose a threat: inverting them can yield photorealistic surrogates that compromise privacy and enable impersonation. Altho…

Face Recognition

Semantic Composition with PSHRG for Derivation Tree Reconstruction from Graph-Based Meaning Representations

2022-05-01 · ACL 2022 5 · Chun Hei Lo, Wai Lam, Hong Cheng

We introduce a data-driven approach to generating derivation trees from meaning representation graphs with probabilistic synchronous hyperedge replacement grammar (PSHRG). SHRG has been used to produce meaning representa…

Semantic Composition

AlignGS: Aligning Geometry and Semantics for Robust Indoor Reconstruction from Sparse Views

2025-10-09 · Yijie Gao, Houqiang Zhong, Tianchi Zhu, Zhengxue Cheng 외 arxiv

The demand for semantically rich 3D models of indoor scenes is rapidly growing, driven by applications in augmented reality, virtual reality, and robotics. However, creating them from sparse views remains a challenge due…

Novel View Synthesis

SoS: Analysis of Surface over Semantics in Multilingual Text-To-Image Generation

2026-01-23 · Carolin Holtermann, Florian Schneider, Anne Lauscher arxiv

Text-to-image (T2I) models are increasingly employed by users worldwide. However, prior research has pointed to the high sensitivity of T2I towards particular input languages - when faced with languages other than Englis…

Text-to-Image Generation

EEG-Driven Image Reconstruction with Saliency-Guided Diffusion Models

2025-10-30 · Igor Abramov, Ilya Makarov arxiv

Existing EEG-driven image reconstruction methods often overlook spatial attention mechanisms, limiting fidelity and semantic coherence. To address this, we propose a dual-conditioning framework that combines EEG embeddin…

Image ReconstructionImage Generation