paper-with-me

Papers

Text-based Animatable 3D Avatars with Morphable Model Alignment

2025-04-22 · Yiqian Wu, Malte Prinzler, Xiaogang Jin, Siyu Tang

The generation of high-quality, animatable 3D head avatars from text has enormous potential in content creation applications such as games, movies, and embodied virtual assistants. Current text-to-3D generation methods typically combine parametric head models with 2D diffusion models using score distillation sampling to produce 3D-consistent results. However, they struggle to synthesize realistic details and suffer from misalignments between the appearance and the driving parametric model, resulting in unnatural animation results. We discovered that these limitations stem from ambiguities in the 2D diffusion predictions during 3D avatar distillation, specifically: i) the avatar's appearance and geometry is underconstrained by the text input, and ii) the semantic alignment between the predictions and the parametric head model is insufficient because the diffusion model alone cannot incorporate information from the parametric model. In this work, we propose a novel framework, AnimPortrait3D, for text-based realistic animatable 3DGS avatar generation with morphable model alignment, and introduce two key strategies to address these challenges. First, we tackle appearance and geometry ambiguities by utilizing prior information from a pretrained text-to-3D model to initialize a 3D avatar with robust appearance, geometry, and rigging relationships to the morphable model. Second, we refine the initial 3D avatar for dynamic expressions using a ControlNet that is conditioned on semantic and normal maps of the morphable model to ensure accurate alignment. As a result, our method outperforms existing approaches in terms of synthesis quality, alignment, and animation fidelity. Our experiments show that the proposed method advances the state of the art in text-based, animatable 3D head avatar generation.

📄 PDF Abstract BibTeX arXiv:2504.15835

Code (1)

onethousand1000/animportrait3d 공식 구현 jax

Tasks

3D Generation3DGSText to 3D

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

High-Fidelity Human Avatars from Laptop Webcams using Edge Compute

2025-02-04 · Akash Haridas, Imran N. Junejo

Applications of generating photo-realistic human avatars are many, however, high-fidelity avatar generation traditionally required expensive professional camera rigs and artistic labor, but recent research has enabled co…

CAP4D: Creating Animatable 4D Portrait Avatars with Morphable Multi-View Diffusion Models

2024-12-16 · CVPR 2025 1 · Felix Taubner, Ruihang Zhang, Mathieu Tuli, David B. Lindell

Reconstructing photorealistic and dynamic portrait avatars from images is essential to many applications including advertising, visual effects, and virtual reality. Depending on the application, avatar reconstruction inv…

Neural Rendering

Morphable Diffusion: 3D-Consistent Diffusion for Single-image Avatar Creation

2024-01-09 · CVPR 2024 1 · Xiyi Chen, Marko Mihajlovic, Shaofei Wang, Sergey Prokudin 외

Recent advances in generative diffusion models have enabled the previously unfeasible capability of generating 3D assets from a single input image or a text prompt. In this work, we aim to enhance the quality and functio…

Novel View Synthesis

PointAvatar: Deformable Point-based Head Avatars from Videos

2022-12-16 · CVPR 2023 1 · Yufeng Zheng, Wang Yifan, Gordon Wetzstein, Michael J. Black 외

The ability to create realistic, animatable and relightable head avatars from casual video sequences would open up wide ranging applications in communication and entertainment. Current methods either build on explicit 3D…

Single-Shot Implicit Morphable Faces with Consistent Texture Parameterization

2023-05-04 · Connor Z. Lin, Koki Nagano, Jan Kautz, Eric R. Chan 외

There is a growing demand for the accessible creation of high-quality 3D avatars that are animatable and customizable. Although 3D morphable models provide intuitive control for editing and animation, and robustness for …

Face ModelFace Reconstruction