paper-with-me

홈 › Papers

ZeroAvatar: Zero-shot 3D Avatar Generation from a Single Image

2023-05-25 · Zhenzhen Weng, Zeyu Wang, Serena Yeung

Recent advancements in text-to-image generation have enabled significant progress in zero-shot 3D shape generation. This is achieved by score distillation, a methodology that uses pre-trained text-to-image diffusion models to optimize the parameters of a 3D neural presentation, e.g. Neural Radiance Field (NeRF). While showing promising results, existing methods are often not able to preserve the geometry of complex shapes, such as human bodies. To address this challenge, we present ZeroAvatar, a method that introduces the explicit 3D human body prior to the optimization process. Specifically, we first estimate and refine the parameters of a parametric human body from a single image. Then during optimization, we use the posed parametric body as additional geometry constraint to regularize the diffusion model as well as the underlying density field. Lastly, we propose a UV-guided texture regularization term to further guide the completion of texture on invisible body parts. We show that ZeroAvatar significantly enhances the robustness and 3D consistency of optimization-based image-to-3D avatar generation, outperforming existing zero-shot image-to-3D methods.

📄 PDF Abstract BibTeX arXiv:2305.16411

Code (0)

등록된 구현이 없습니다.

Tasks

3D Shape GenerationImage GenerationImage to 3DNeRFText to Image GenerationText-to-Image Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

GAIA: Zero-shot Talking Avatar Generation

2023-11-26 · Tianyu He, Junliang Guo, Runyi Yu, Yuchi Wang 외

Zero-shot talking avatar generation aims at synthesizing natural talking videos from speech and a single portrait image. Previous methods have relied on domain-specific heuristics such as warping-based motion representat…

Diversity

AvatarCLIP: Zero-Shot Text-Driven Generation and Animation of 3D Avatars

2022-05-17 · Fangzhou Hong, Mingyuan Zhang, Liang Pan, Zhongang Cai 외

3D avatar creation plays a crucial role in the digital age. However, the whole production process is prohibitively time-consuming and labor-intensive. To democratize this technology to a larger audience, we propose Avata…

3D geometryLanguage ModellingMotion SynthesisTexture Synthesis

AvatarFusion: Zero-shot Generation of Clothing-Decoupled 3D Avatars Using 2D Diffusion

2023-07-13 · Shuo Huang, Zongxin Yang, Liangting Li, Yi Yang 외

Large-scale pre-trained vision-language models allow for the zero-shot text-based generation of 3D avatars. The previous state-of-the-art method utilized CLIP to supervise neural implicit models that reconstructed a huma…

Text-Conditional Contextualized Avatars For Zero-Shot Personalization

2023-04-14 · Samaneh Azadi, Thomas Hayes, Akbar Shah, Guan Pang 외

Recent large-scale text-to-image generation models have made significant improvements in the quality, realism, and diversity of the synthesized images and enable users to control the created content through language. How…

DiversityImage GenerationText to 3DText to Image Generation+1

Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion

2025-03-20 · CVPR 2025 1 · Zhou Zhenglin, Ma Fan, Fan Hehe, Chua Tat-Seng

Animatable head avatar generation typically requires extensive data for training. To reduce the data requirements, a natural solution is to leverage existing data-free static avatar generation methods, such as pre-traine…