paper-with-me

Papers

RevealNet: Seeing Behind Objects in RGB-D Scans

2019-04-26 · CVPR 2020 6 · Ji Hou, Angela Dai, Matthias Nießner

During 3D reconstruction, it is often the case that people cannot scan each individual object from all views, resulting in missing geometry in the captured scan. This missing geometry can be fundamentally limiting for many applications, e.g., a robot needs to know the unseen geometry to perform a precise grasp on an object. Thus, we introduce the task of semantic instance completion: from an incomplete RGB-D scan of a scene, we aim to detect the individual object instances and infer their complete object geometry. This will open up new possibilities for interactions with objects in a scene, for instance for virtual or robotic agents. We tackle this problem by introducing RevealNet, a new data-driven approach that jointly detects object instances and predicts their complete geometry. This enables a semantically meaningful decomposition of a scanned scene into individual, complete 3D objects, including hidden and unobserved object parts. RevealNet is an end-to-end 3D neural network architecture that leverages joint color and geometry feature learning. The fully-convolutional nature of our 3D network enables efficient inference of semantic instance completion for 3D scans at scale of large indoor environments in a single forward pass. We show that predicting complete object geometry improves both 3D detection and instance segmentation performance. We evaluate on both real and synthetic scan benchmark data for the new task, where we outperform state-of-the-art approaches by over 15 in mAP@0.5 on ScanNet, and over 18 in mAP@0.5 on SUNCG.

📄 PDF Abstract BibTeX arXiv:1904.12012

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction3D Semantic Instance SegmentationInstance SegmentationObjectSemantic Segmentation

Similar Papers 제목 키워드 기반

Seeing the Wind from a Falling Leaf

2025-11-30 · Zhiyuan Gao, Jiageng Mao, Hong-Xing Yu, Haozhe Lou 외 arxiv

A longstanding goal in computer vision is to model motions from videos, while the representations behind motions, i.e. the invisible physical interactions that cause objects to deform and move, remain largely unexplored.…

Video Generation

Seeing and Seeing Through the Glass: Real and Synthetic Data for Multi-Layer Depth Estimation

2025-03-14 · Hongyu Wen, Yiming Zuo, Venkat Subramanian, Patrick Chen 외

Transparent objects are common in daily life, and understanding their multi-layer depth information -- perceiving both the transparent surface and the objects behind it -- is crucial for real-world applications that inte…

Depth EstimationTransparent objects

Learning to Grasp Without Seeing

2018-05-10 · Adithyavairavan Murali, Yin Li, Dhiraj Gandhi, Abhinav Gupta

Can a robot grasp an unknown object without seeing it? In this paper, we present a tactile-sensing based approach to this challenging problem of grasping novel objects without prior knowledge of their location or physica…

Object Localization

Variational Amodal Object Completion

2020-12-01 · NeurIPS 2020 12 · Huan Ling, David Acuna, Karsten Kreis, Seung Wook Kim 외

In images of complex scenes, objects are often occluding each other which makes perception tasks such as object detection and tracking, or robotic control tasks such as planning, challenging. To facilitate downstream tas…

Objectobject-detectionObject Detection

Seeing Behind Objects for 3D Multi-Object Tracking in RGB-D Sequences

2020-12-15 · CVPR 2021 1 · Norman Müller, Yu-Shiang Wong, Niloy J. Mitra, Angela Dai 외

Multi-object tracking from RGB-D video sequences is a challenging problem due to the combination of changing viewpoints, motion, and occlusions over time. We observe that having the complete geometry of objects aids in t…

3D Multi-Object TrackingMulti-Object TrackingObjectObject Tracking