paper-with-me

홈 › Papers

One Map to Find Them All: Real-time Open-Vocabulary Mapping for Zero-shot Multi-Object Navigation

2024-09-18 · Finn Lukas Busch, Timon Homberger, Jesús Ortega-Peimbert, Quantao Yang, Olov Andersson

The capability to efficiently search for objects in complex environments is fundamental for many real-world robot applications. Recent advances in open-vocabulary vision models have resulted in semantically-informed object navigation methods that allow a robot to search for an arbitrary object without prior training. However, these zero-shot methods have so far treated the environment as unknown for each consecutive query. In this paper we introduce a new benchmark for zero-shot multi-object navigation, allowing the robot to leverage information gathered from previous searches to more efficiently find new objects. To address this problem we build a reusable open-vocabulary feature map tailored for real-time object search. We further propose a probabilistic-semantic map update that mitigates common sources of errors in semantic feature extraction and leverage this semantic uncertainty for informed multi-object exploration. We evaluate our method on a set of object navigation tasks in both simulation as well as with a real robot, running in real-time on a Jetson Orin AGX. We demonstrate that it outperforms existing state-of-the-art approaches both on single and multi-object navigation tasks. Additional videos, code and the multi-object navigation benchmark will be available on https://finnbsch.github.io/OneMap.

📄 PDF Abstract BibTeX arXiv:2409.11764

Code (0)

등록된 구현이 없습니다.

Tasks

AllObject

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

OTAS: Open-vocabulary Token Alignment for Outdoor Segmentation

2025-07-08 · Simon Schwaiger, Stefan Thalhammer, Wilfried Wöber, Gerald Steinbauer-Wagner arxiv

Understanding open-world semantics is critical for robotic planning and control, particularly in unstructured outdoor environments. Existing vision-language mapping approaches typically rely on object-centric segmentatio…

HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation

2024-09-22 · Naoki Yokoyama, Ram Ramrakhya, Abhishek Das, Dhruv Batra 외

We present the Habitat-Matterport 3D Open Vocabulary Object Goal Navigation dataset (HM3D-OVON), a large-scale benchmark that broadens the scope and semantic range of prior Object Goal Navigation (ObjectNav) benchmarks. …

NavigateVisual Navigation

FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment

2025-04-11 · Sebastián Barbas Laina, Simon Boche, Sotiris Papatheodorou, Simon Schaefer 외

Geometrically accurate and semantically expressive map representations have proven invaluable to facilitate robust and safe mobile robot navigation and task planning. Nevertheless, real-time, open-vocabulary semantic und…

3D geometryNatural Language QueriesRobot NavigationScene Understanding+1

Open-YOLO 3D: Towards Fast and Accurate Open-Vocabulary 3D Instance Segmentation

2024-06-04 · Mohamed El Amine Boudjoghra, Angela Dai, Jean Lahoud, Hisham Cholakkal 외

Recent works on open-vocabulary 3D instance segmentation show strong promise, but at the cost of slow inference speed and high computation requirements. This high computation cost is typically due to their heavy reliance…

2D Object Detection3D Instance Segmentation3D Open-Vocabulary Instance SegmentationInstance Segmentation+4

OCTO+: A Suite for Automatic Open-Vocabulary Object Placement in Mixed Reality

2024-01-17 · Aditya Sharma, Luke Yoffe, Tobias Höllerer

One key challenge in Augmented Reality is the placement of virtual content in natural locations. Most existing automated techniques can only work with a closed-vocabulary, fixed set of objects. In this paper, we introduc…

Mixed Realityvalid