paper-with-me

홈 › Papers

ReScene4D: Temporally Consistent Semantic Instance Segmentation of Evolving Indoor 3D Scenes

2026-01-16 · Emily Steiner, Jianhao Zheng, Henry Howard-Jenkins, Chris Xie, Iro Armeni arxiv

Indoor environments evolve as objects move, appear, or leave the scene. Capturing these dynamics requires maintaining temporally consistent instance identities across intermittently captured 3D scans, even when changes are unobserved. We introduce and formalize the task of temporally sparse 4D indoor semantic instance segmentation (SIS), which jointly segments, identifies, and temporally associates object instances. This setting poses a challenge for existing 3DSIS methods, which require a discrete matching step due to their lack of temporal reasoning, and for 4D LiDAR approaches, which perform poorly due to their reliance on high-frequency temporal measurements that are uncommon in the longer-horizon evolution of indoor environments. We propose ReScene4D, a novel method that adapts 3DSIS architectures for 4DSIS without needing dense observations. Our method enables temporal information sharing--using spatiotemporal contrastive loss, masking, and serialization--to adaptively leverage geometric and semantic priors across observations. This shared context enables consistent instance tracking and improves standard 3DSIS performance. To evaluate this task, we define a new metric, t-mAP, that extends mAP to reward temporal identity consistency. ReScene4D achieves state-of-the-art performance on the 3RScan dataset, establishing a new benchmark for understanding evolving indoor scenes.

📄 PDF Abstract BibTeX arXiv:2601.11508

Code (0)

등록된 구현이 없습니다.

Tasks

Instance Segmentation

Similar Papers 제목 키워드 기반

Semantically Coherent Co-Segmentation and Reconstruction of Dynamic Scenes

2017-07-01 · CVPR 2017 7 · Armin Mustafa, Adrian Hilton

In this paper we propose a framework for spatially and temporally coherent semantic co-segmentation and reconstruction of complex dynamic scenes from multiple static or moving cameras. Semantic co-segmentation exploits t…

3D ReconstructionSegmentation

ReScene: Structured Indoor Scene Reconstruction from Multi-View Captures

2026-06-26 · Haoran Xu, Lechao Zhang, Daoguo Dong, Yan Gao 외 arxiv

Constructing simulation-ready 3D scenes from multi-view captures is a key bottleneck for Embodied Artificial Intelligence, as downstream tasks require object-level structure, explicit inter-object relations, and physical…

Visual Question AnsweringSpatial Reasoning

A Benchmark for LiDAR-based Panoptic Segmentation based on KITTI

2020-03-04 · Jens Behley, Andres Milioto, Cyrill Stachniss

Panoptic segmentation is the recently introduced task that tackles semantic segmentation and instance segmentation jointly. In this paper, we present an extension of SemanticKITTI, which is a large-scale dataset providin…

Instance SegmentationPanoptic SegmentationSegmentationSemantic Segmentation

4D Panoptic LiDAR Segmentation

2021-02-24 · CVPR 2021 1 · Mehmet Aygün, Aljoša Ošep, Mark Weber, Maxim Maximov 외

Temporal semantic scene understanding is critical for self-driving cars or robots operating in dynamic environments. In this paper, we propose 4D panoptic LiDAR segmentation to assign a semantic class and a temporally-co…

4D Panoptic SegmentationBenchmarkingMulti-Object TrackingObject Tracking+3

U4D: Unsupervised 4D Dynamic Scene Understanding

2019-07-23 · ICCV 2019 10 · Armin Mustafa, Chris Russell, Adrian Hilton

We introduce the first approach to solve the challenging problem of unsupervised 4D visual scene understanding for complex dynamic scenes with multiple interacting people from multi-view video. Our approach simultaneousl…

3D Pose EstimationInstance SegmentationPose EstimationScene Understanding+2