paper-with-me

홈 › Papers

MOHO: Learning Single-view Hand-held Object Reconstruction with Multi-view Occlusion-Aware Supervision

2023-10-18 · CVPR 2024 1 · Chenyangguang Zhang, Guanlong Jiao, Yan Di, Gu Wang, Ziqin Huang, Ruida Zhang, Fabian Manhardt, Bowen Fu, Federico Tombari, Xiangyang Ji

Previous works concerning single-view hand-held object reconstruction typically rely on supervision from 3D ground-truth models, which are hard to collect in real world. In contrast, readily accessible hand-object videos offer a promising training data source, but they only give heavily occluded object observations. In this paper, we present a novel synthetic-to-real framework to exploit Multi-view Occlusion-aware supervision from hand-object videos for Hand-held Object reconstruction (MOHO) from a single image, tackling two predominant challenges in such setting: hand-induced occlusion and object's self-occlusion. First, in the synthetic pre-training stage, we render a large-scaled synthetic dataset SOMVideo with hand-object images and multi-view occlusion-free supervisions, adopted to address hand-induced occlusion in both 2D and 3D spaces. Second, in the real-world finetuning stage, MOHO leverages the amodal-mask-weighted geometric supervision to mitigate the unfaithful guidance caused by the hand-occluded supervising views in real world. Moreover, domain-consistent occlusion-aware features are amalgamated in MOHO to resist object's self-occlusion for inferring the complete object shape. Extensive experiments on HO3D and DexYCB datasets demonstrate 2D-supervised MOHO gains superior results against 3D-supervised methods by a large margin.

📄 PDF Abstract BibTeX arXiv:2310.11696

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectObject Reconstruction

Similar Papers 제목 키워드 기반

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis

2026-07-19 · Lingwei Dang, Juntong Li, Zonghan Li, Hongwen Zhang 외 hf

Hand-Object Interaction (HOI) synthesis is a cornerstone for animation production and embodied AI. Despite the strong priors of video foundation models, multi-view consistent HOI synthesis remains challenging due to comp…

Contact-conditioned hand-held object reconstruction from single-view images

2023-08-01 · journal 2023 8 · Xiaoyuan Wang, Yang Li, Adnane Boukhayma, Changbo Wang 외

Reconstructing the shape of hand-held objects from single-view color images is a long-standing problem in computer vision and computer graphics. The task is complicated by the ill-posed nature of single-view reconstructi…

ObjectObject Reconstruction

NVS-HO: A Benchmark for Novel View Synthesis of Handheld Objects

2026-02-05 · Musawar Ali, Manuel Carranza-García, Nicola Fioraio, Samuele Salti 외 arxiv

We propose NVS-HO, the first benchmark designed for novel view synthesis of handheld objects in real-world environments using only RGB inputs. Each object is recorded in two complementary RGB sequences: (1) a handheld se…

Novel View Synthesis

3D Reconstruction of Objects in Hands without Real World 3D Supervision

2023-05-04 · Aditya Prakash, Matthew Chang, Matthew Jin, Ruisen Tu 외

Prior works for reconstructing hand-held objects from a single image train models on images paired with 3D shapes. Such data is challenging to gather in the real world at scale. Consequently, these approaches do not gene…

3D ReconstructionObjectObject Reconstruction

Towards Scalable Multi-View Reconstruction of Geometry and Materials

2023-06-06 · Carolin Schmitt, Božidar Antić, Andrei Neculai, Joo Ho Lee 외

In this paper, we propose a novel method for joint recovery of camera pose, object geometry and spatially-varying Bidirectional Reflectance Distribution Function (svBRDF) of 3D scenes that exceed object-scale and hence c…

Distributed Optimization