paper-with-me

Papers

Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation

2026-05-24 · Aviral Chharia, Fernando De la Torre arxiv

High-fidelity 3D Gaussian head avatar generation is critical for applications such as AR/VR, telepresence, and digital humans. Existing methods depend on multi-view datasets, 3D captures, or intermediate 2D view synthesis. In contrast, we learn both conditional and unconditional 3D head models from randomly sampled 2D images alone, without using multi-view data, 3D supervision, or intermediate view generation. We introduce MVCHead, a single-shot state space model that enforces multi-view consistency (MVC) directly in the 3D representation while regressing 3D Gaussians under these constraints. At its core, we propose a Hierarchical State Space (HiSS) block that progressively refines Gaussians from coarse to fine, while capturing long-range dependencies. Within each HiSS block, we modify Mamba's standard unidirectional scan with the proposed Hierarchical Bi-directional State Scan (HiBiSS) that aligns recurrence with the axes along which multi-view inconsistencies are strongest. Finally, we design an SE(3) Multi-view Critic that judges whether a set of self-renders arises from a single underlying 3D configuration, rewarding cross-view pixel alignment without observing real multi-view pairs. MVCHead achieves state-of-the-art perceptual quality, surpasses prior methods in both texture and geometric consistency, and maintains comparable shape consistency. To demonstrate scalability, we release FaceGS-10K, the first large-scale dataset of ready-to-use 3D Gaussian head assets for training and evaluation of 3D head models. Project Page and code: https://humansensinglab.github.io/MVCHead/

📄 PDF Abstract BibTeX arXiv:2605.25220

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AvatarBack: Back-Head Generation for Complete 3D Avatars from Front-View Images

2025-08-28 · Shiqi Xin, Xiaolin Zhang, Yanbin Liu, Peng Zhang 외 arxiv

Recent advances in Gaussian Splatting have significantly boosted the reconstruction of head avatars, enabling high-quality facial modeling by representing an 3D avatar as a collection of 3D Gaussians. However, existing m…

HeadStudio: Text to Animatable Head Avatars with 3D Gaussian Splatting

2024-02-09 · Zhenglin Zhou, Fan Ma, Hehe Fan, Zongxin Yang 외

Creating digital avatars from textual prompts has long been a desirable yet challenging task. Despite the promising results achieved with 2D diffusion priors, current methods struggle to create high-quality and consisten…

Generating Editable Head Avatars with 3D Gaussian GANs

2024-12-26 · Guohao Li, Hongyu Yang, Yifang Men, Di Huang 외

Generating animatable and editable 3D head avatars is essential for various applications in computer vision and graphics. Traditional 3D-aware generative adversarial networks (GANs), often using implicit fields like Neur…

3DGSNeRF

One-Shot Feed-Forward 360$^{\circ}$ Animatable Avatar via Inpainted UV-Space Gaussian Modeling

2026-01-19 · Shuling Zhao, Dan Xu arxiv

Building one-shot 3D animatable head avatars is an important yet challenging problem. Existing methods generally collapse under large camera pose variations, compromising the realism of 3D avatars. In this work, we propo…

3D Reconstruction

RelightAnyone: A Generalized Relightable 3D Gaussian Head Model

2026-01-06 · Yingyan Xu, Pramod Rao, Sebastian Weiss, Gaspard Zoss 외 arxiv

3D Gaussian Splatting (3DGS) has become a standard approach to reconstruct and render photorealistic 3D head avatars. A major challenge is to relight the avatars to match any scene illumination. For high quality relighti…

Novel View Synthesis