paper-with-me

Papers

MVP4D: Multi-View Portrait Video Diffusion for Animatable 4D Avatars

2025-10-14 · Felix Taubner, Ruihang Zhang, Mathieu Tuli, Sherwin Bahmani, David B. Lindell arxiv

Digital human avatars aim to simulate the dynamic appearance of humans in virtual environments, enabling immersive experiences across gaming, film, virtual reality, and more. However, the conventional process for creating and animating photorealistic human avatars is expensive and time-consuming, requiring large camera capture rigs and significant manual effort from professional 3D artists. With the advent of capable image and video generation models, recent methods enable automatic rendering of realistic animated avatars from a single casually captured reference image of a target subject. While these techniques significantly lower barriers to avatar creation and offer compelling realism, they lack constraints provided by multi-view information or an explicit 3D representation. So, image quality and realism degrade when rendered from viewpoints that deviate strongly from the reference image. Here, we build a video model that generates animatable multi-view videos of digital humans based on a single reference image and target expressions. Our model, MVP4D, is based on a state-of-the-art pre-trained video diffusion model and generates hundreds of frames simultaneously from viewpoints varying by up to 360 degrees around a target subject. We show how to distill the outputs of this model into a 4D avatar that can be rendered in real-time. Our approach significantly improves the realism, temporal consistency, and 3D consistency of generated avatars compared to previous methods.

📄 PDF Abstract BibTeX arXiv:2510.12785

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

CosAvatar: Consistent and Animatable Portrait Video Tuning with Text Prompt

2023-11-30 · Haiyao Xiao, Chenglai Zhong, Xuan Gao, Yudong Guo 외

Recently, text-guided digital portrait editing has attracted more and more attentions. However, existing methods still struggle to maintain consistency across time, expression, and view or require specific data prerequis…

NeRF

SOAP: Style-Omniscient Animatable Portraits

2025-05-08 · Tingting Liao, Yujian Zheng, Adilbek Karmanov, Liwen Hu 외

Creating animatable 3D avatars from a single image remains challenging due to style limitations (realistic, cartoon, anime) and difficulties in handling accessories or hairstyles. While 3D diffusion models advance single…

Image to 3D

CAP4D: Creating Animatable 4D Portrait Avatars with Morphable Multi-View Diffusion Models

2024-12-16 · CVPR 2025 1 · Felix Taubner, Ruihang Zhang, Mathieu Tuli, David B. Lindell

Reconstructing photorealistic and dynamic portrait avatars from images is essential to many applications including advertising, visual effects, and virtual reality. Depending on the application, avatar reconstruction inv…

Neural Rendering

AniPortraitGAN: Animatable 3D Portrait Generation from 2D Image Collections

2023-09-05 · Yue Wu, Sicheng Xu, Jianfeng Xiang, Fangyun Wei 외

Previous animatable 3D-aware GANs for human generation have primarily focused on either the human head or full body. However, head-only videos are relatively uncommon in real life, and full body generation typically does…

Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos

2025-08-12 · Chaoyi Wang, Yifan Yang, Jun Pei, Lijie Xia 외 arxiv

Creating realistic, fully animatable whole-body avatars from a single portrait is challenging due to limitations in capturing subtle expressions, body movements, and dynamic backgrounds. Current evaluation datasets and m…