paper-with-me

홈 › Papers

SpatialAvatar-0: High-Quality 4D Head Avatar with Multi-Stage Reconstruction

2026-06-14 · Yiran Wang, Zeyu Zhang, Yuanming Li, Ziming Wang, Yang Zhao arxiv

High-quality 4D head avatars from one or a few source portraits are central to telepresence, AR/VR, and digital-human interaction. 3D Gaussian Splatting (3DGS) has emerged as the dominant representation, with two complementary regimes (generalizable feed-forward predictors and per-subject refiners) maturing in parallel. However, existing feed-forward predictors are trained on a single dataset family with a hard-coded source count, inheriting the corresponding domain bias. Per-subject refiners require 300K--600K iterations and rely on adaptive densification that destroys upstream Gaussian layouts, preventing the two regimes from sharing a representation end-to-end. To bridge both regimes we propose SpatialAvatar-0 on a shared FLAME-mesh-bound Gaussian representation: a feed-forward generator with a parameter-free K-source mean-pool and a monocular-temporal to multi-view-spatial two-phase schedule that anchors against identity-prior collapse onto the smaller multi-view set. We further introduce a 10K-iter layout-preserving per-subject refinement loop that freezes the FLAME-binding and Gaussian count and replaces densification with a three-component anti-spike regularization. On VFHQ/HDTF cross-domain zero-shot we surpass the in-domain leader GAGAvatar by +1.5 dB PSNR despite never training on either test domain, and on the SplattingAvatar monocular benchmark we lead every reported metric, surpassing the 300K-iter GeoAvatar by +1.3 dB PSNR at up to 60x shorter per-subject schedule than common SOTA baselines. Website: https://spatialwalk.github.io/SpatialAvatar-0.

📄 PDF Abstract BibTeX arXiv:2606.15659

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

HRAvatar: High-Quality and Relightable Gaussian Head Avatar

2025-03-11 · CVPR 2025 1 · Dongbin Zhang, Yunfei Liu, Lijian Lin, Ye Zhu 외

Reconstructing animatable and high-quality 3D head avatars from monocular videos, especially with realistic relighting, is a valuable task. However, the limited information from single-view input, combined with the compl…

3DGS

One2Avatar: Generative Implicit Head Avatar For Few-shot User Adaptation

2024-02-19 · Zhixuan Yu, Ziqian Bai, Abhimitra Meka, Feitong Tan 외

Traditional methods for constructing high-quality, personalized head avatars from monocular videos demand extensive face captures and training time, posing a significant challenge for scalability. This paper introduces a…

Camera Calibration

HeadStudio: Text to Animatable Head Avatars with 3D Gaussian Splatting

2024-02-09 · Zhenglin Zhou, Fan Ma, Hehe Fan, Zongxin Yang 외

Creating digital avatars from textual prompts has long been a desirable yet challenging task. Despite the promising results achieved with 2D diffusion priors, current methods struggle to create high-quality and consisten…

Any3DAvatar: Fast and High-Quality Full-Head 3D Avatar Reconstruction from Single Portrait Image

2026-04-15 · Yujie Gao, Yao Xiao, Xiangnan Zhu, Ya Li 외 arxiv

Reconstructing a complete 3D head from a single portrait remains challenging because existing methods still face a sharp quality-speed trade-off: high-fidelity pipelines often rely on multi-stage processing and per-subje…

OPHAvatars: One-shot Photo-realistic Head Avatars

2023-07-18 · Shaoxu Li

We propose a method for synthesizing photo-realistic digital avatars from only one portrait as the reference. Given a portrait, our method synthesizes a coarse talking head video using driving keypoints features. And wit…

Blind Face Restoration