paper-with-me

Papers

InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction from Cluttered Scenes

2025-07-11 · Zesong Yang, Bangbang Yang, Wenqi Dong, Chenxuan Cao, Liyuan Cui, Yuewen Ma, Zhaopeng Cui, Hujun Bao arxiv

Humans can naturally identify and mentally complete occluded objects in cluttered environments. However, imparting similar cognitive ability to robotics remains challenging even with advanced reconstruction techniques, which models scenes as undifferentiated wholes and fails to recognize complete object from partial observations. In this paper, we propose InstaScene, a new paradigm towards holistic 3D perception of complex scenes with a primary goal: decomposing arbitrary instances while ensuring complete reconstruction. To achieve precise decomposition, we develop a novel spatial contrastive learning by tracing rasterization of each instance across views, significantly enhancing semantic supervision in cluttered scenes. To overcome incompleteness from limited observations, we introduce in-situ generation that harnesses valuable observations and geometric cues, effectively guiding 3D generative models to reconstruct complete instances that seamlessly align with the real world. Experiments on scene decomposition and object completion across complex real-world and synthetic scenes demonstrate that our method achieves superior decomposition accuracy while producing geometrically faithful and visually intact objects.

📄 PDF Abstract BibTeX arXiv:2507.08416

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learning

Similar Papers 제목 키워드 기반

From Local Matches to Global Masks: Template-Guided Instance Detection and Segmentation in Open-World Scenes

2026-03-03 · Qifan Zhang, Sai Haneesh Allu, Jikai Wang, Yangxiao Lu 외 arxiv

Detecting and segmenting novel object instances in open-world environments is a fundamental problem in robotic perception. Given only a small set of template images, a robot must locate and segment a specific object inst…

Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling

2026-08-31 · Minghan Qin, Yuang Wang, Xiuyu Yang, Yushi Long 외 hf

Composable scene modeling aims to recover a real indoor scene as complete, editable object assets arranged as observed, giving robot simulation and embodied AI a simulation-ready replica of the real environment whose obj…

3D Object DetectionPose Estimation

Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes

2026-02-09 · Seunghoon Jeong, Eunho Lee, Jeongyun Kim, Ayoung Kim arxiv

In cluttered scenes with inevitable occlusions and incomplete observations, selecting informative viewpoints is essential for building a reliable representation. In this context, 3D Gaussian Splatting (3DGS) offers a dis…

RevealNet: Seeing Behind Objects in RGB-D Scans

2019-04-26 · CVPR 2020 6 · Ji Hou, Angela Dai, Matthias Nießner

During 3D reconstruction, it is often the case that people cannot scan each individual object from all views, resulting in missing geometry in the captured scan. This missing geometry can be fundamentally limiting for ma…

3D Reconstruction3D Semantic Instance SegmentationInstance SegmentationObject+1

Fusion++: Volumetric Object-Level SLAM

2018-08-25 · John McCormac, Ronald Clark, Michael Bloesch, Andrew J. Davison 외

We propose an online object-level SLAM system which builds a persistent and accurate 3D graph map of arbitrary reconstructed objects. As an RGB-D camera browses a cluttered indoor scene, Mask-RCNN instance segmentations …

Loop Closure DetectionObject