paper-with-me

홈 › Papers

Revealing Directions for Text-guided 3D Face Editing

2024-10-07 · Zhuo Chen, Yichao Yan, Sehngqi Liu, Yuhao Cheng, Weiming Zhao, Lincheng Li, Mengxiao Bi, Xiaokang Yang

3D face editing is a significant task in multimedia, aimed at the manipulation of 3D face models across various control signals. The success of 3D-aware GAN provides expressive 3D models learned from 2D single-view images only, encouraging researchers to discover semantic editing directions in its latent space. However, previous methods face challenges in balancing quality, efficiency, and generalization. To solve the problem, we explore the possibility of introducing the strength of diffusion model into 3D-aware GANs. In this paper, we present Face Clan, a fast and text-general approach for generating and manipulating 3D faces based on arbitrary attribute descriptions. To achieve disentangled editing, we propose to diffuse on the latent space under a pair of opposite prompts to estimate the mask indicating the region of interest on latent codes. Based on the mask, we then apply denoising to the masked latent codes to reveal the editing direction. Our method offers a precisely controllable manipulation method, allowing users to intuitively customize regions of interest with the text description. Experiments demonstrate the effectiveness and generalization of our Face Clan for various pre-trained GANs. It offers an intuitive and wide application for text-guided face editing that contributes to the landscape of multimedia content creation.

📄 PDF Abstract BibTeX arXiv:2410.04965

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeDenoising

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

CLIP-Guided StyleGAN Inversion for Text-Driven Real Image Editing

2023-07-17 · Ahmet Canberk Baykal, Abdul Basit Anees, Duygu Ceylan, Erkut Erdem 외

Researchers have recently begun exploring the use of StyleGAN-based models for real image editing. One particularly interesting application is using natural language descriptions to guide the editing process. Existing ap…

Attribute

DeltaSpace: A Semantic-aligned Feature Space for Flexible Text-guided Image Editing

2023-10-12 · Yueming Lyu, Kang Zhao, Bo Peng, Yue Jiang 외

Text-guided image editing faces significant challenges to training and inference flexibility. Much literature collects large amounts of annotated image-text pairs to train text-conditioned generative models from scratch,…

text-guided-image-editing

Text-Guided 3D Face Synthesis -- From Generation to Editing

2023-12-01 · Yunjie Wu, Yapeng Meng, Zhipeng Hu, Lincheng Li 외

Text-guided 3D face synthesis has achieved remarkable results by leveraging text-to-image (T2I) diffusion models. However, most existing works focus solely on the direct generation, ignoring the editing, restricting them…

Face GenerationTexture Synthesis

Text-Guided 3D Face Synthesis - From Generation to Editing

2024-01-01 · CVPR 2024 1 · Yunjie Wu, Yapeng Meng, Zhipeng Hu, Lincheng Li 외

Text-guided 3D face synthesis has achieved remarkable results by leveraging text-to-image (T2I) diffusion models. However most existing works focus solely on the direct generation ignoring the editing restricting the…

Face GenerationTexture Synthesis

A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models

2024-06-20 · Xincheng Shuai, Henghui Ding, Xingjun Ma, RongCheng Tu 외

Image editing aims to edit the given synthetic or real image to meet the specific requirements from users. It is widely studied in recent years as a promising and challenging field of Artificial Intelligence Generative C…

Video Editing