paper-with-me

Papers

RoboAtlas: Contextual Active SLAM

2026-06-24 · Alexander Schperberg, Shivam K. Panda, Abraham P. Vinod, M. K. Jawed, Stefano Di Cairano arxiv

We present RoboAtlas, a contextual Active SLAM framework that adaptively balances geometric exploration and semantic reasoning using a scalable 3D semantic mapping system, OpenRoboVox. RoboAtlas integrates frontier exploration, global semantic-map reasoning, and egocentric VLM-based reasoning through a contextual multi-armed bandit that transitions from exploration to semantically guided navigation as scene understanding improves. We evaluate the system in simulation and on a Unitree Go2 robot in large-scale real-world environments exceeding 1800 m2 with approx. 30k mapped semantic instances, achieving a 100% task success rate. On the GOAT-Bench "Val Unseen" benchmark, RoboAtlas achieves state-of-the-art performance with highest reported success rate (SR) of 90.6%, using GPT-4o, improving over the strongest prior baseline by 17.8 percentage points in SR. Using the much smaller Qwen2.5-VL-7B model, it still achieves 88.8% SR, outperforming all baselines using GPT-4o in SR, and revealing the importance of the information gained by our semantic mapping framework over simply replacing the underlying foundation model. The results demonstrate that grounding foundation models with large-scale 3D semantic maps enables robust and efficient contextual Active SLAM.

📄 PDF Abstract BibTeX arXiv:2606.26046

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Understanding

Similar Papers 제목 키워드 기반

Implementing a Sharia Chatbot as a Consultation Medium for Questions About Islam

2025-12-18 · Wisnu Uriawan, Aria Octavian Hamza, Ade Ripaldi Nuralim, Adi Purnama 외 arxiv

This research presents the implementation of a Sharia-compliant chatbot as an interactive medium for consulting Islamic questions, leveraging Reinforcement Learning (Q-Learning) integrated with Sentence-Transformers for …

Reinforcement Learning

Dream-SLAM: Dreaming the Unseen for Active SLAM in Dynamic Environments

2026-02-25 · Xiangqi Meng, Pengxu Hou, Zhenjun Zhao, Javier Civera 외 arxiv

In addition to the core tasks of simultaneous localization and mapping (SLAM), active SLAM additionally in- volves generating robot actions that enable effective and efficient exploration of unknown environments. However…

Camera Pose EstimationMotion Planning

Stereo 3D Gaussian Splatting SLAM for Outdoor Urban Scenes

2025-07-31 · Xiaohan Li, Ziren Gong, Fabio Tosi, Matteo Poggi 외 arxiv

3D Gaussian Splatting (3DGS) has recently gained popularity in SLAM applications due to its fast rendering and high-fidelity representation. However, existing 3DGS-SLAM systems have predominantly focused on indoor enviro…

A Novel ViDAR Device With Visual Inertial Encoder Odometry and Reinforcement Learning-Based Active SLAM Method

2025-06-16 · Zhanhua Xin, Zhihao Wang, Shenghao Zhang, Wanchao Chi 외

In the field of multi-sensor fusion for simultaneous localization and mapping (SLAM), monocular cameras and IMUs are widely used to build simple and effective visual-inertial systems. However, limited research has explor…

Deep Reinforcement LearningSensor FusionSimultaneous Localization and MappingState Estimation

Dropping the D: RGB-D SLAM Without the Depth Sensor

2025-10-07 · Mert Kiray, Alican Karaomer, Benjamin Busam arxiv

We present DropD-SLAM, a real-time monocular SLAM system that achieves RGB-D-level accuracy without relying on depth sensors. The system replaces active depth input with three pretrained vision modules: a monocular metri…

Instance Segmentation