paper-with-me

Papers

CosAvatar: Consistent and Animatable Portrait Video Tuning with Text Prompt

2023-11-30 · Haiyao Xiao, Chenglai Zhong, Xuan Gao, Yudong Guo, Juyong Zhang

Recently, text-guided digital portrait editing has attracted more and more attentions. However, existing methods still struggle to maintain consistency across time, expression, and view or require specific data prerequisites. To solve these challenging problems, we propose CosAvatar, a high-quality and user-friendly framework for portrait tuning. With only monocular video and text instructions as input, we can produce animatable portraits with both temporal and 3D consistency. Different from methods that directly edit in the 2D domain, we employ a dynamic NeRF-based 3D portrait representation to model both the head and torso. We alternate between editing the video frames' dataset and updating the underlying 3D portrait until the edited frames reach 3D consistency. Additionally, we integrate the semantic portrait priors to enhance the edited results, allowing precise modifications in specified semantic areas. Extensive results demonstrate that our proposed method can not only accurately edit portrait styles or local attributes based on text instructions but also support expressive animation driven by a source video.

📄 PDF Abstract BibTeX arXiv:2311.18288

Code (0)

등록된 구현이 없습니다.

Tasks

NeRF

Similar Papers 제목 키워드 기반

AniPortraitGAN: Animatable 3D Portrait Generation from 2D Image Collections

2023-09-05 · Yue Wu, Sicheng Xu, Jianfeng Xiang, Fangyun Wei 외

Previous animatable 3D-aware GANs for human generation have primarily focused on either the human head or full body. However, head-only videos are relatively uncommon in real life, and full body generation typically does…

PERSE: Personalized 3D Generative Avatars from A Single Portrait

2024-12-30 · CVPR 2025 1 · Hyunsoo Cha, Inhee Lee, Hanbyul Joo

We present PERSE, a method for building an animatable personalized generative avatar from a reference portrait. Our avatar model enables facial attribute editing in a continuous and disentangled latent space to control e…

Attribute

OPHAvatars: One-shot Photo-realistic Head Avatars

2023-07-18 · Shaoxu Li

We propose a method for synthesizing photo-realistic digital avatars from only one portrait as the reference. Given a portrait, our method synthesizes a coarse talking head video using driving keypoints features. And wit…

Blind Face Restoration

GeoDiff4D: Geometry-Aware Diffusion for 4D Head Avatar Reconstruction

2026-02-27 · Chao Xu, Xiaochen Zhao, Xiang Deng, Jingxiang Sun 외 arxiv

Reconstructing photorealistic and animatable 4D head avatars from a single portrait image remains a fundamental challenge in computer vision. While diffusion models have enabled remarkable progress in image and video gen…

Video Generation

Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos

2025-08-12 · Chaoyi Wang, Yifan Yang, Jun Pei, Lijie Xia 외 arxiv

Creating realistic, fully animatable whole-body avatars from a single portrait is challenging due to limitations in capturing subtle expressions, body movements, and dynamic backgrounds. Current evaluation datasets and m…