paper-with-me

Papers

WildAvatar: Learning In-the-wild 3D Avatars from the Web

2025-01-01 · CVPR 2025 1 · Zihao Huang, Shoukang Hu, Guangcong Wang, Tianqi Liu, Yuhang Zang, Zhiguo Cao, Wei Li, Ziwei Liu

Existing research on avatar creation is typically limited to laboratory datasets, which require high costs against scalability and exhibit insufficient representation of the real world. On the other hand, the web abounds with off-the-shelf real-world human videos, but these videos vary in quality and require accurate annotations for avatar creation. To this end, we propose an automatic annotating pipeline with filtering protocols to curate these humans from the web. Our pipeline surpasses state-of-the-art methods on the EMDB benchmark, and the filtering protocols boost verification metrics on web videos. We then curate WildAvatar, a web-scale in-the-wild human avatar creation dataset extracted from YouTube, with 10,000+ different human subjects and scenes. WildAvatar is at least 10xricher than previous datasets for 3D human avatar creation and closer to the real world. To explore its potential, we demonstrate the quality and generalizability of avatar creation methods on WildAvatar. We will publicly release our code, data source links and annotations to push forward 3D human avatar creation and other related fields for real-world applications.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

WildAvatar: Web-scale In-the-wild Video Dataset for 3D Avatar Creation

2024-07-02 · Zihao Huang, Shoukang Hu, Guangcong Wang, Tianqi Liu 외

Existing human datasets for avatar creation are typically limited to laboratory environments, wherein high-quality annotations (e.g., SMPL estimation from 3D scans or multi-view images) can be ideally provided. However, …

L3D-Pose: Lifting Pose for 3D Avatars from a Single Camera in the Wild

2025-01-02 · Soumyaratna Debnath, Harish Katti, Shashikant Verma, Shanmuganathan Raman

While 2D pose estimation has advanced our ability to interpret body movements in animals and primates, it is limited by the lack of depth information, constraining its application range. 3D pose estimation provides a mor…

2D Pose Estimation3D Pose EstimationPose Estimation

WildGHand: Learning Anti-Perturbation Gaussian Hand Avatars from Monocular In-the-Wild Videos

2026-02-24 · Hanhui Li, Xuan Huang, Wanquan Liu, Yuhao Cheng 외 arxiv

Despite recent progress in 3D hand reconstruction from monocular videos, most existing methods rely on data captured in well-controlled environments and therefore degrade in real-world settings with severe perturbations,…

Realistic One-shot Mesh-based Head Avatars

2022-06-16 · Taras Khakhulin, Vanessa Sklyarova, Victor Lempitsky, Egor Zakharov

We present a system for realistic one-shot mesh-based human head avatars creation, ROME for short. Using a single photograph, our model estimates a person-specific head mesh and the associated neural texture, which encod…

FAGhead: Fully Animate Gaussian Head from Monocular Videos

2024-06-27 · Yixin Xuan, Xinyang Li, Gongxin Yao, Shiwei Zhou 외

High-fidelity reconstruction of 3D human avatars has a wild application in visual reality. In this paper, we introduce FAGhead, a method that enables fully controllable human portraits from monocular videos. We explicit …