paper-with-me

홈 › Papers

AHOY! Animatable Humans under Occlusion from YouTube Videos with Gaussian Splatting and Video Diffusion Priors

2026-03-18 · Aymen Mir, Riza Alp Guler, Xiangjun Tang, Peter Wonka, Gerard Pons-Moll arxiv

We present AHOY, a method for reconstructing complete, animatable 3D Gaussian avatars from in-the-wild monocular video despite heavy occlusion. Existing methods assume unoccluded input-a fully visible subject, often in a canonical pose-excluding the vast majority of real-world footage where people are routinely occluded by furniture, objects, or other people. Reconstructing from such footage poses fundamental challenges: large body regions may never be observed, and multi-view supervision per pose is unavailable. We address these challenges with four contributions: (i) a hallucination-as-supervision pipeline that uses identity-finetuned diffusion models to generate dense supervision for previously unobserved body regions; (ii) a two-stage canonical-to-pose-dependent architecture that bootstraps from sparse observations to full pose-dependent Gaussian maps; (iii) a map-pose/LBS-pose decoupling that absorbs multi-view inconsistencies from the generated data; (iv) a head/body split supervision strategy that preserves facial identity. We evaluate on YouTube videos and on multi-view capture data with significant occlusion and demonstrate state-of-the-art reconstruction quality. We also demonstrate that the resulting avatars are robust enough to be animated with novel poses and composited into 3DGS scenes captured using cell-phone video. Our project page is available at https://miraymen.github.io/ahoy/

📄 PDF Abstract BibTeX arXiv:2603.17975

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ARCH: Animatable Reconstruction of Clothed Humans

2020-04-08 · CVPR 2020 6 · Zeng Huang, Yuanlu Xu, Christoph Lassner, Hao Li 외

In this paper, we propose ARCH (Animatable Reconstruction of Clothed Humans), a novel end-to-end framework for accurate reconstruction of animation-ready 3D clothed humans from a monocular image. Existing approaches to d…

3D Object Reconstruction From A Single Image3D Reconstruction

Relightable and Animatable Neural Avatars from Videos

2023-12-20 · Wenbin Lin, Chengwei Zheng, Jun-Hai Yong, Feng Xu

Lightweight creation of 3D digital avatars is a highly desirable but challenging task. With only sparse videos of a person under unknown illumination, we propose a method to create relightable and animatable neural avata…

InpaintHuman: Reconstructing Occluded Humans with Multi-Scale UV Mapping and Identity-Preserving Diffusion Inpainting

2026-01-05 · Jinlong Fan, Shanshan Zhao, Liang Zheng, Jing Zhang 외 arxiv

Reconstructing complete and animatable 3D human avatars from monocular videos remains challenging, particularly under severe occlusions. While 3D Gaussian Splatting has enabled photorealistic human rendering, existing me…

iHuman: Instant Animatable Digital Humans From Monocular Videos

2024-07-15 · Pramish Paudel, Anubhav Khanal, Ajad Chhatkuli, Danda Pani Paudel 외

Personalized 3D avatars require an animatable representation of digital humans. Doing so instantly from monocular videos offers scalability to broad class of users and wide-scale applications. In this paper, we present a…

3D geometry3D Reconstruction

GASPACHO: Gaussian Splatting for Controllable Humans and Objects

2025-03-12 · Aymen Mir, Arthur Moreau, Helisa Dhamo, Zhensong Zhang 외

We present GASPACHO: a method for generating photorealistic controllable renderings of human-object interactions. Given a set of multi-view RGB images of human-object interactions, our method reconstructs animatable temp…

Human-Object Interaction DetectionObject