Contrastive Instance Association for 4D Panoptic Segmentation using Sequences of 3D LiDAR Scans
Scene understanding is critical for autonomous navigation in dynamic environments. Perception tasks in this domain like segmentation and tracking are usually tackled individually. In this paper, we address the problem of 4D panoptic segmentation using LiDAR scans, which requires to assign to each 3D point in a temporal sequence of scans a semantic class and for each object a temporally consistent instance ID. We propose a novel approach that builds on top of an arbitrary single-scan panoptic segmentation network and extends it to the temporal domain by associating instances across time. We propose a contrastive aggregation network that leverages the point-wise features from the panoptic network. It generates an embedding space in which encodings of the same instance at different timesteps lie close together and far from encodings belonging to other instances. The training is inspired by contrastive learning techniques for self-supervised metric learning. Our association module combines appearance and motion cues to associate instances across scans, allowing us to perform temporal perception. We evaluate our proposed method on the SemanticKITTI benchmark and achieve state-of-the-art results even without relying on pose information.
Code (1)
Tasks
4D Panoptic SegmentationAutonomous NavigationContrastive LearningMetric LearningPanoptic SegmentationScene UnderstandingSegmentationSimilar Papers 제목 키워드 기반
Mask4D: End-to-End Mask-Based 4D Panoptic Segmentation for LiDAR Sequences
Scene understanding is crucial for autonomous systems to reliably navigate in the real world. Panoptic segmentation of 3D LiDAR scans allows us to semantically describe a vehicle’s environment by predicting semantic clas…
3D Panoptic Segmentation4D Panoptic SegmentationNavigatePanoptic Segmentation+2PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
Panoptic reconstruction is a challenging task in 3D scene understanding. However, most existing methods heavily rely on pre-trained semantic segmentation models and known 3D object bounding boxes for 3D panoptic segmenta…
3D Panoptic SegmentationInstance SegmentationPanoptic SegmentationScene Understanding+3Open-Set LiDAR Panoptic Segmentation Guided by Uncertainty-Aware Learning
Autonomous vehicles that navigate in open-world environments may encounter previously unseen object classes. However, most existing LiDAR panoptic segmentation models rely on closed-set assumptions, failing to detect unk…
Autonomous VehiclesNavigatePanoptic SegmentationSegmentation+1Mask4Former: Mask Transformer for 4D Panoptic Segmentation
Accurately perceiving and tracking instances over time is essential for the decision-making processes of autonomous agents interacting safely in dynamic environments. With this intention, we propose Mask4Former for the c…
4D Panoptic SegmentationInstance SegmentationObject TrackingPanoptic Segmentation+14D Panoptic LiDAR Segmentation
Temporal semantic scene understanding is critical for self-driving cars or robots operating in dynamic environments. In this paper, we propose 4D panoptic LiDAR segmentation to assign a semantic class and a temporally-co…
4D Panoptic SegmentationBenchmarkingMulti-Object TrackingObject Tracking+3