paper-with-me

Papers

Reconstructing Humans and Objects in Interaction using Large Reconstruction Models

2026-08-27 · Agniv Chatterjee, Georgios Pavlakos arxiv

Estimation of Human-Object Interactions in 3D (3D HOI) is a fundamental problem in 3D computer vision with applications in AR/VR, robotics, and embodied AI. However, reconstructing these interactions in 3D remains challenging due to depth ambiguities, occlusions, and object shape variability. Existing approaches are primarily concerned with reprojection and contact constraints, fitting parametric human models and object templates to 2D images. In this paper, we explore a different avenue. We present MILO, a framework that leverages the visual capabilities of Large Reconstruction Models (LRMs) to recover detailed 3D human-object interactions from a single image. Our key observation is that LRMs provide a powerful geometric scaffold that preserves relative human-object arrangement and proximity cues. This significantly simplifies the reconstruction procedure, reframing the problem as interpreting the LRM mesh: we segment it into human and object components, fit a parametric body model to the human part, and optionally align an object template to the object part (if such a template is available). MILO achieves strong reconstruction accuracy and outperforms existing baselines across multiple benchmarks and interaction scenarios. Our code is available at https://ac5113.github.io/MILO.

📄 PDF Abstract BibTeX arXiv:2608.27407

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions

2025-03-20 · CVPR 2025 1 · Boran Wen, Dingbang Huang, Zichen Zhang, Jiahong Zhou 외

Reconstructing human-object interactions (HOI) from single images is fundamental in computer vision. Existing methods are primarily trained and tested on indoor scenes due to the lack of 3D data, particularly constrained…

3D ReconstructionHuman-Object Interaction Detection

Single-image coherent reconstruction of objects and humans

2024-08-15 · Sarthak Batra, Partha P. Chakrabarti, Simon Hadfield, Armin Mustafa

Existing methods for reconstructing objects and humans from a monocular image suffer from severe mesh collisions and performance limitations for interacting occluding objects. This paper introduces a method to obtain a g…

3D ReconstructionImage Inpainting

CHORD: Category-level Hand-held Object Reconstruction via Shape Deformation

2023-08-21 · ICCV 2023 1 · Kailin Li, Lixin Yang, Haoyu Zhen, Zenan Lin 외

In daily life, humans utilize hands to manipulate objects. Modeling the shape of objects that are manipulated by the hand is essential for AI to comprehend daily tasks and to learn manipulation skills. However, previous …

Object Reconstruction

RHINO: Reconstructing Human Interactions with Novel Objects from Monocular Videos

2026-05-16 · Lixin Xue, Chengwei Zheng, Georgios Paschalidis, Chen Guo 외 arxiv

Reconstructing people, objects, and their interactions in 3D is a long-standing goal for intelligent systems. Often the input is RGB video from a moving camera, making the task ill-posed; depth is ambiguous, humans and o…

Reconstructing Action-Conditioned Human-Object Interactions Using Commonsense Knowledge Priors

2022-09-06 · Xi Wang, Gen Li, Yen-Ling Kuo, Muhammed Kocabas 외

We present a method for inferring diverse 3D models of human-object interactions from images. Reasoning about how humans interact with objects in complex scenes from a single 2D image is a challenging task given ambiguit…

Human-Object Interaction DetectionObject