paper-with-me

홈 › Papers

Progressive Pose-Guided 4D Animal Reconstruction from Monocular Video

2026-06-30 · Siyuan Li, Weiying Chen, Yilin Wang, Xinxin Zuo, Xingyu Li, Li Cheng arxiv

Reconstructing 4D animals from monocular videos is challenging due to large inter-species variation, complex articulations, and the lack of reliable templates. Existing approaches typically rely on either strict category-specific priors that restrict generalization, or unconstrained generative models that sacrifice input fidelity. To bridge this gap, we present a progressive test-time optimization framework built on 3D Gaussian Splatting for high-fidelity 4D animal reconstruction from a single video. Our key insight is that a coarse shape prior suffices when coupled with a progressive strategy that disentangles articulated pose from non-rigid deformation. Specifically, we employ a symmetry-aware temporal encoding that exploits bilateral cues while absorbing camera estimation drift and a part-conditioned deformation mechanism guided by learnable part anchors and a learnable skinning field. Extensive experiments demonstrate that our approach generalizes robustly across diverse species, achieving superior geometric accuracy, temporal consistency, and visual fidelity compared to existing baselines, even under severe prior mismatch.

📄 PDF Abstract BibTeX arXiv:2607.00157

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction

2026-06-11 · Jen-Hao Cheng, Yipeng Wang, Hao Zhang, Gengshan Yang 외 arxiv

We present Flex4DHuman, a multi-view video diffusion model that transforms a monocular or sparse multi-view video of a dynamic subject into synchronized dense multi-view videos using only relative camera-pose conditionin…

DensePose 3D: Lifting Canonical Surface Maps of Articulated Objects to the Third Dimension

2021-08-31 · ICCV 2021 10 · Roman Shapovalov, David Novotny, Benjamin Graham, Patrick Labatut 외

We tackle the problem of monocular 3D reconstruction of articulated objects like humans and animals. We contribute DensePose 3D, a method that can learn such reconstructions in a weakly supervised fashion from 2D image a…

3D ReconstructionMonocular ReconstructionObject

PPR: Physically Plausible Reconstruction from Monocular Videos

2023-01-01 · ICCV 2023 1 · Gengshan Yang, Shuo Yang, John Z. Zhang, Zachary Manchester 외

Given monocular videos, we build 3D models of articulated objects and environments whose 3D configurations satisfy dynamics and contact constraints. At its core, our method leverages differentiable physics simulation…

CASA: Category-agnostic Skeletal Animal Reconstruction

2022-11-04 · Yuefan Wu, Zeyuan Chen, Shaowei Liu, Zhongzheng Ren 외

Recovering the skeletal shape of an animal from a monocular video is a longstanding challenge. Prevailing animal reconstruction methods often adopt a control-point driven animation model and optimize bone transforms indi…

Retrieval

Recovering Physically Plausible Human-Object Interactions from Monocular Videos

2026-06-03 · Dingbang Huang, Etienne Vouga, Qixing Huang, Georgios Pavlakos arxiv

In this paper, we propose RePHO, a method to reconstruct physically plausible human-object interactions (HOI) from monocular videos. While existing kinematic-based approaches produce visually plausible motion, they often…

Reinforcement Learning