paper-with-me

Papers

WildAvatar: Web-scale In-the-wild Video Dataset for 3D Avatar Creation

2024-07-02 · Zihao Huang, Shoukang Hu, Guangcong Wang, Tianqi Liu, Yuhang Zang, Zhiguo Cao, Wei Li, Ziwei Liu

Existing human datasets for avatar creation are typically limited to laboratory environments, wherein high-quality annotations (e.g., SMPL estimation from 3D scans or multi-view images) can be ideally provided. However, their annotating requirements are impractical for real-world images or videos, posing challenges toward real-world applications on current avatar creation methods. To this end, we propose the WildAvatar dataset, a web-scale in-the-wild human avatar creation dataset extracted from YouTube, with $10,000+$ different human subjects and scenes. WildAvatar is at least $10\times$ richer than previous datasets for 3D human avatar creation. We evaluate several state-of-the-art avatar creation methods on our dataset, highlighting the unexplored challenges in real-world applications on avatar creation. We also demonstrate the potential for generalizability of avatar creation methods, when provided with data at scale. We publicly release our data source links and annotations, to push forward 3D human avatar creation and other related fields for real-world applications.

📄 PDF Abstract BibTeX arXiv:2407.02165

Code (1)

wildavatar/WildAvatar_Toolbox 공식 구현 pytorch

Similar Papers 제목 키워드 기반

WildAvatar: Learning In-the-wild 3D Avatars from the Web

2025-01-01 · CVPR 2025 1 · Zihao Huang, Shoukang Hu, Guangcong Wang, Tianqi Liu 외

Existing research on avatar creation is typically limited to laboratory datasets, which require high costs against scalability and exhibit insufficient representation of the real world. On the other hand, the web abo…

Vid2Avatar-Pro: Authentic Avatar from Videos in the Wild via Universal Prior

2025-03-03 · CVPR 2025 1 · Chen Guo, Junxuan Li, Yash Kant, Yaser Sheikh 외

We present Vid2Avatar-Pro, a method to create photorealistic and animatable 3D human avatars from monocular in-the-wild videos. Building a high-quality avatar that supports animation with diverse poses from a monocular v…

Inverse Rendering

WildGHand: Learning Anti-Perturbation Gaussian Hand Avatars from Monocular In-the-Wild Videos

2026-02-24 · Hanhui Li, Xuan Huang, Wanquan Liu, Yuhao Cheng 외 arxiv

Despite recent progress in 3D hand reconstruction from monocular videos, most existing methods rely on data captured in well-controlled environments and therefore degrade in real-world settings with severe perturbations,…

Large-scale Codec Avatars: The Unreasonable Effectiveness of Large-scale Avatar Pretraining

2026-04-02 · Junxuan Li, Rawal Khirodkar, Chengan He, Zhongshi Jiang 외 arxiv

High-quality 3D avatar modeling faces a critical trade-off between fidelity and generalization. On the one hand, multi-view studio data enables high-fidelity modeling of humans with precise control over expressions and p…

GenLCA: 3D Diffusion for Full-Body Avatars from In-the-Wild Videos

2026-04-08 · Yiqian Wu, Rawal Khirodkar, Egor Zakharov, Timur Bagautdinov 외 arxiv

We present GenLCA, a diffusion-based generative model for generating and editing photorealistic full-body avatars from text and image inputs. The generated avatars are faithful to the inputs, while supporting high-fideli…