paper-with-me

홈 › Papers

Hand-held Object Reconstruction from RGB Video with Dynamic Interaction

2025-01-01 · CVPR 2025 1 · Shijian Jiang, Qi Ye, Rengan Xie, Yuchi Huo, Jiming Chen

This work aims to reconstruct the 3D geometry of a rigid object manipulated by one or both hands using monocular RGB video. Previous methods rely on Structure-from-Motion or hand priors to estimate relative motion between the object and camera, which typically assume textured objects or single-hand interactions. To accurately recover object geometry in dynamic interactions, we incorporate priors from 3D generation model into object pose estimation and propose semantic consistency constraints to solve the challenge of shape and texture discrepancy between the generated priors and observations. The poses are initialized, followed by joint optimization of the object poses and implicit neural representation. During optimization, a novel pose outlier voting strategy with inter-view consistency is proposed to correct large pose errors. Experiments on three datasets demonstrate that our method significantly outperforms the state-of-the-art in reconstruction quality for both single- and two-hand scenarios. Our project page: https://east-j.github.io/dynhor/

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

3D Generation3D geometryObjectObject ReconstructionPose Estimation

Similar Papers 제목 키워드 기반

MOHO: Learning Single-view Hand-held Object Reconstruction with Multi-view Occlusion-Aware Supervision

2023-10-18 · CVPR 2024 1 · Chenyangguang Zhang, Guanlong Jiao, Yan Di, Gu Wang 외

Previous works concerning single-view hand-held object reconstruction typically rely on supervision from 3D ground-truth models, which are hard to collect in real world. In contrast, readily accessible hand-object videos…

ObjectObject Reconstruction

Reconstructing Hand-Held Objects from Monocular Video

2022-11-30 · Di Huang, Xiaopeng Ji, Xingyi He, Jiaming Sun 외

This paper presents an approach that reconstructs a hand-held object from a monocular video. In contrast to many recent methods that directly predict object geometry by a trained network, the proposed approach does not r…

Hand Pose EstimationObjectPose Estimation

Reconstructing Hand-Held Objects in 3D from Images and Videos

2024-04-09 · Jane Wu, Georgios Pavlakos, Georgia Gkioxari, Jitendra Malik

Objects manipulated by the hand (i.e., manipulanda) are particularly challenging to reconstruct from Internet videos. Not only does the hand occlude much of the object, but also the object is often only visible in a smal…

ObjectObject ReconstructionText to 3D

Generalizable Articulated Object Reconstruction from Casually Captured RGBD Videos

2025-06-10 · Weikun Peng, Jun Lv, Cewu Lu, Manolis Savva

Articulated objects are prevalent in daily life. Understanding their kinematic structure and reconstructing them have numerous applications in embodied AI and robotics. However, current methods require carefully captured…

ObjectObject Reconstruction

Learning the Depths of Moving People by Watching Frozen People

2019-04-25 · CVPR 2019 6 · Zhengqi Li, Tali Dekel, Forrester Cole, Richard Tucker 외

We present a method for predicting dense depth in scenarios where both a monocular camera and people in the scene are freely moving. Existing methods for recovering depth for dynamic, non-rigid objects from monocular vid…

Depth EstimationDepth Prediction