paper-with-me

홈 › Papers

Better Together: Unified Motion Capture and 3D Avatar Reconstruction

2025-03-12 · Arthur Moreau, Mohammed Brahimi, Richard Shaw, Athanasios Papaioannou, Thomas Tanay, Zhensong Zhang, Eduardo Pérez-Pellitero

We present Better Together, a method that simultaneously solves the human pose estimation problem while reconstructing a photorealistic 3D human avatar from multi-view videos. While prior art usually solves these problems separately, we argue that joint optimization of skeletal motion with a 3D renderable body model brings synergistic effects, i.e. yields more precise motion capture and improved visual quality of real-time rendering of avatars. To achieve this, we introduce a novel animatable avatar with 3D Gaussians rigged on a personalized mesh and propose to optimize the motion sequence with time-dependent MLPs that provide accurate and temporally consistent pose estimates. We first evaluate our method on highly challenging yoga poses and demonstrate state-of-the-art accuracy on multi-view human pose estimation, reducing error by 35% on body joints and 45% on hand joints compared to keypoint-based methods. At the same time, our method significantly boosts the visual quality of animatable avatars (+2dB PSNR on novel view synthesis) on diverse challenging subjects.

📄 PDF Abstract BibTeX arXiv:2503.09293

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View SynthesisPose Estimation

Similar Papers 제목 키워드 기반

UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars

2026-05-14 · Xiaoyu Zhan, Xinyu Fu, Chenghao Yang, Xiaohong Zhang 외 arxiv

Speech-driven gestures and facial animations are fundamental to expressive digital avatars in games, virtual production, and interactive media. However, existing methods are either limited to a single modality for audio …

UNICA: A Unified Neural Framework for Controllable 3D Avatars

2026-04-03 · Jiahe Zhu, Xinyao Wang, Yiyu Zhuang, Yanwen Wang 외 arxiv

Controllable 3D human avatars have found widespread applications in 3D games, the metaverse, and AR/VR scenarios. The conventional approach to creating such a 3D avatar requires a lengthy, intricate pipeline encompassing…

Motion Planning

EgoAvatar: Egocentric View-Driven and Photorealistic Full-body Avatars

2024-09-22 · Jianchun Chen, Jian Wang, yinda zhang, Rohit Pandey 외

Immersive VR telepresence ideally means being able to interact and communicate with digital avatars that are indistinguishable from and precisely reflect the behaviour of their real counterparts. The core technical chall…

Structure-Aware Fine-Grained Gaussian Splatting for Expressive Avatar Reconstruction

2026-04-10 · Yuze Su, Hongsong Wang, Jie Gui, Liang Wang arxiv

Reconstructing photorealistic and topology-aware human avatars from monocular videos remains a significant challenge in the fields of computer vision and graphics. While existing 3D human avatar modeling approaches can e…

HumanSplatHMR: Closing the Loop Between Human Mesh Recovery and Gaussian Splatting Avatar

2026-05-04 · Yeheng Zong, Pou-Chun Kung, Yike Pan, Seth Isaacson 외 arxiv

Accurately recovering human pose and appearance from video is an essential component of scene reconstruction, with applications to motion capture, motion prediction, virtual reality, and digital twinning. Despite signifi…

Human Mesh RecoveryPose Estimation