Mapping the Unseen: Unified Promptable Panoptic Mapping with Dynamic Labeling using Foundation Models
In the field of robotics and computer vision, efficient and accurate semantic mapping remains a significant challenge due to the growing demand for intelligent machines that can comprehend and interact with complex environments. Conventional panoptic mapping methods, however, are limited by predefined semantic classes, thus making them ineffective for handling novel or unforeseen objects. In response to this limitation, we introduce the Unified Promptable Panoptic Mapping (UPPM) method. UPPM utilizes recent advances in foundation models to enable real-time, on-demand label generation using natural language prompts. By incorporating a dynamic labeling strategy into traditional panoptic mapping techniques, UPPM provides significant improvements in adaptability and versatility while maintaining high performance levels in map reconstruction. We demonstrate our approach on real-world and simulated datasets. Results show that UPPM can accurately reconstruct scenes and segment objects while generating rich semantic labels through natural language interactions. A series of ablation experiments validated the advantages of foundation model-based labeling over fixed label sets.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
PanopticNDT: Efficient and Robust Panoptic Mapping
As the application scenarios of mobile robots are getting more complex and challenging, scene understanding becomes increasingly crucial. A mobile robot that is supposed to operate autonomously in indoor environments mus…
2D Panoptic Segmentation3D Panoptic Segmentation3D Semantic SegmentationPanoptic Segmentation+4A Review of Panoptic Segmentation for Mobile Mapping Point Clouds
3D point cloud panoptic segmentation is the combined task to (i) assign each point to a semantic class and (ii) separate the points in each class into object instances. Recently there has been an increased interest in su…
Instance SegmentationPanoptic SegmentationScene UnderstandingSegmentation+1PanoSLAM: Panoptic 3D Scene Reconstruction via Gaussian SLAM
Understanding geometric, semantic, and instance information in 3D scenes from sequential video data is essential for applications in robotics and augmented reality. However, existing Simultaneous Localization and Mapping…
3D Instance Segmentation3D Reconstruction3D Scene Reconstruction3D Semantic Segmentation+5PanopticFusion: Online Volumetric Semantic Mapping at the Level of Stuff and Things
We propose PanopticFusion, a novel online volumetric semantic mapping system at the level of stuff and things. In contrast to previous semantic mapping systems, PanopticFusion is able to densely predict class labels of a…
3D Instance SegmentationInstance SegmentationPanoptic SegmentationSemantic SegmentationPhenoStitch: Training-Free Panoptic Crop Mapping from Satellite Image Time Series
Panoptic crop mapping requires both delineating individual agricultural parcels and assigning a crop type to each parcel from satellite image time series. Existing approaches typically rely on dense parcel-level annotati…