paper-with-me

Papers

CompSLAM: Complementary Hierarchical Multi-Modal Localization and Mapping for Robot Autonomy in Underground Environments

2025-05-10 · Shehryar Khattak, Timon Homberger, Lukas Bernreiter, Julian Nubert, Olov Andersson, Roland Siegwart, Kostas Alexis, Marco Hutter

Robot autonomy in unknown, GPS-denied, and complex underground environments requires real-time, robust, and accurate onboard pose estimation and mapping for reliable operations. This becomes particularly challenging in perception-degraded subterranean conditions under harsh environmental factors, including darkness, dust, and geometrically self-similar structures. This paper details CompSLAM, a highly resilient and hierarchical multi-modal localization and mapping framework designed to address these challenges. Its flexible architecture achieves resilience through redundancy by leveraging the complementary nature of pose estimates derived from diverse sensor modalities. Developed during the DARPA Subterranean Challenge, CompSLAM was successfully deployed on all aerial, legged, and wheeled robots of Team Cerberus during their competition-winning final run. Furthermore, it has proven to be a reliable odometry and mapping solution in various subsequent projects, with extensions enabling multi-robot map sharing for marsupial robotic deployments and collaborative mapping. This paper also introduces a comprehensive dataset acquired by a manually teleoperated quadrupedal robot, covering a significant portion of the DARPA Subterranean Challenge finals course. This dataset evaluates CompSLAM's robustness to sensor degradations as the robot traverses 740 meters in an environment characterized by highly variable geometries and demanding lighting conditions. The CompSLAM code and the DARPA SubT Finals dataset are made publicly available for the benefit of the robotics community

📄 PDF Abstract BibTeX arXiv:2505.06483

Code (1)

jokerjohn/cloud_map_evaluation

Tasks

Pose Estimation

Similar Papers 제목 키워드 기반

Hierarchical Multi-Process Fusion for Visual Place Recognition

2020-01-28 · Stephen Hausler, Michael Milford

Combining multiple complementary techniques together has long been regarded as a way to improve performance. In visual localization, multi-sensor fusion, multi-process fusion of a single sensing modality, and even combin…

Sensor FusionVisual LocalizationVisual Place Recognition

FusionViT: Hierarchical 3D Object Detection via LiDAR-Camera Vision Transformer Fusion

2023-11-07 · Xinhao Xiang, Jiawei Zhang

For 3D object detection, both camera and lidar have been demonstrated to be useful sensory devices for providing complementary information about the same scenery with data representations in different modalities, e.g., 2…

3D Object DetectionObjectobject-detectionObject Detection+2

Localizing Audio-Visual Deepfakes via Hierarchical Boundary Modeling

2025-08-04 · Xuanjun Chen, Shih-Peng Cheng, Jiawei Du, Lin Zhang 외 arxiv

Audio-visual temporal deepfake localization under the content-driven partial manipulation remains a highly challenging task. In this scenario, the deepfake regions are usually only spanning a few frames, with the majorit…

Multi-GAT: A Graphical Attention-based Hierarchical Multimodal Representation Learning Approach for Human Activity Recognition

2021-04-01 · IEEE ROBOTICS AND AUTOMATION LETTERS 2021 4 · Md Mofijul Islam, Tariq Iqbal

Recognizing human activities is one of the crucial capabilities that a robot needs to have to be useful around people. Although modern robots are equipped with various types of sensors, human activity recognition (HAR) s…

Activity RecognitionHuman Activity RecognitionMixture-of-ExpertsMultimodal Activity Recognition+1

GLEAM: A Multimodal Imaging Dataset and HAMM for Glaucoma Classification

2026-03-13 · Jiao Wang, Chi Liu, Yiying Zhang, Hongchen Luo 외 arxiv

We propose glaucoma lesion evaluation and analysis with multimodal imaging (GLEAM), the first publicly available tri-modal glaucoma dataset comprising scanning laser ophthalmoscopy fundus images, circumpapillary OCT imag…

Representation Learning