paper-with-me

Papers

Point2Graph: An End-to-end Point Cloud-based 3D Open-Vocabulary Scene Graph for Robot Navigation

2024-09-16 · Yifan Xu, Ziming Luo, Qianwei Wang, Vineet Kamat, Carol Menassa

Current open-vocabulary scene graph generation algorithms highly rely on both 3D scene point cloud data and posed RGB-D images and thus have limited applications in scenarios where RGB-D images or camera poses are not readily available. To solve this problem, we propose Point2Graph, a novel end-to-end point cloud-based 3D open-vocabulary scene graph generation framework in which the requirement of posed RGB-D image series is eliminated. This hierarchical framework contains room and object detection/segmentation and open-vocabulary classification. For the room layer, we leverage the advantage of merging the geometry-based border detection algorithm with the learning-based region detection to segment rooms and create a "Snap-Lookup" framework for open-vocabulary room classification. In addition, we create an end-to-end pipeline for the object layer to detect and classify 3D objects based solely on 3D point cloud data. Our evaluation results show that our framework can outperform the current state-of-the-art (SOTA) open-vocabulary object and room segmentation and classification algorithm on widely used real-scene datasets.

📄 PDF Abstract BibTeX arXiv:2409.10350

Code (0)

등록된 구현이 없습니다.

Tasks

3D Open-Vocabulary Object DetectionGraph GenerationObjectobject-detectionObject DetectionRobot NavigationScene Graph Generation

Similar Papers 제목 키워드 기반

Open3DSG: Open-Vocabulary 3D Scene Graphs from Point Clouds with Queryable Objects and Open-Set Relationships

2024-02-19 · CVPR 2024 1 · Sebastian Koch, Narunas Vaskevicius, Mirco Colosi, Pedro Hermosilla 외

Current approaches for 3D scene graph prediction rely on labeled datasets to train models for a fixed set of known object classes and relationship categories. We present Open3DSG, an alternative approach to learn 3D scen…

3d scene graph generationObjectPrediction

Open-Vocabulary 3D Detection via Image-level Class and Debiased Cross-modal Contrastive Learning

2022-07-05 · Yuheng Lu, Chenfeng Xu, Xiaobao Wei, Xiaodong Xie 외

Current point-cloud detection methods have difficulty detecting the open-vocabulary objects in the real world, due to their limited generalization capability. Moreover, it is extremely laborious and expensive to collect …

Cloud DetectionContrastive Learning

HAECcity: Open-Vocabulary Scene Understanding of City-Scale Point Clouds with Superpoint Graph Clustering

2025-04-18 · Alexander Rusnak, Frédéric Kaplan

Traditional 3D scene understanding techniques are generally predicated on hand-annotated label sets, but in recent years a new class of open-vocabulary 3D scene understanding techniques has emerged. Despite the success o…

ClusteringGraph ClusteringMixture-of-ExpertsScene Understanding

Open-Vocabulary Point-Cloud Object Detection without 3D Annotation

2023-04-03 · CVPR 2023 1 · Yuheng Lu, Chenfeng Xu, Xiaobao Wei, Xiaodong Xie 외

The goal of open-vocabulary detection is to identify novel objects based on arbitrary textual descriptions. In this paper, we address open-vocabulary 3D point-cloud detection by a dividing-and-conquering strategy, which …

3D Object Detection3D Open-Vocabulary Object DetectionCloud DetectionContrastive Learning+3

JOPP-3D: Joint Open Vocabulary Semantic Segmentation on Point Clouds and Panoramas

2026-03-06 · Sandeep Inuganti, Hideaki Kanayama, Kanta Shimizu, Mahdi Chamseddine 외 arxiv

Semantic segmentation across visual modalities such as 3D point clouds and panoramic images remains a challenging task, primarily due to the scarcity of annotated data and the limited adaptability of fixed-label models. …

Open Vocabulary Semantic Segmentation3D Semantic SegmentationScene UnderstandingPoint Clouds