paper-with-me

Papers

CAVERS: Multimodal SLAM Data from a Natural Karstic Cave with Ground Truth Motion Capture

2026-04-16 · Giacomo Franchini, David Rodríguez-Martínez, Alfonso Martínez-Petersen, C. J. Pérez-del-Pulgar, Marcello Chiaberge arxiv

Autonomous robots operating in natural karstic caves face perception and navigation challenges that are qualitatively distinct from those encountered in mines or tunnels: irregular geometry, reflective wet surfaces, near-zero ambient light, and complex branching passages. Yet publicly available datasets targeting this environment remain scarce and offer limited sensing modalities and environmental diversity. We present CAVERS, a multimodal dataset acquired in two structurally distinct rooms of Cueva de la Victoria, Málaga, Spain, containing 24 sequences with approximately 335 GB of recorded data. The sensor suite combines an Intel RealSense D435i RGB-D-I camera, an Optris PI640i near-IR thermal camera, and a Velodyne VLP-16 LiDAR, operated both handheld and mounted on a wheeled rover under full darkness and artificial illumination. For most of the sequences, mm-accurate 6-DoF ground truth pose and velocity at 120 Hz are provided by an Optitrack motion capture system installed directly inside the cave. We benchmark seven state-of-the-art SLAM and odometry algorithms spanning visual, visual-inertial, thermal-inertial, and LiDAR-based pipelines, as well as a 3D reconstruction pipeline, demonstrating the dataset's usability. The dataset and all supplementary material are publicly available at: https://github.com/spaceuma/cavers.

📄 PDF Abstract BibTeX arXiv:2604.15052

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction

Similar Papers 제목 키워드 기반

Multimodal Fusion SLAM with Fourier Attention

2025-06-22 · Youjie Zhou, Guofeng Mei, Yiming Wang, Yi Wan 외

Visual SLAM is particularly challenging in environments affected by noise, varying lighting conditions, and darkness. Learning-based optical flow algorithms can leverage multiple modalities to address these challenges, b…

Knowledge DistillationOptical Flow Estimation

Learning to Act with Affordance-Aware Multimodal Neural SLAM

2022-01-24 · Zhiwei Jia, Kaixiang Lin, Yizhou Zhao, Qiaozi Gao 외

Recent years have witnessed an emerging paradigm shift toward embodied artificial intelligence, in which an agent must learn to solve challenging tasks by interacting with its environment. There are several challenges in…

Efficient ExplorationTest unseen

The Rosario Dataset v2: Multimodal Dataset for Agricultural Robotics

2025-08-29 · Nicolas Soncini, Javier Cremona, Erica Vidal, Maximiliano García 외 arxiv

We present a multi-modal dataset collected in a soybean crop field, comprising over two hours of recorded data from sensors such as stereo infrared camera, color camera, accelerometer, gyroscope, magnetometer, GNSS (Sing…

A Hybrid Learner for Simultaneous Localization and Mapping

2021-01-04 · Thangarajah Akilan, Edna Johnson, Japneet Sandhu, Ritika Chadha 외

Simultaneous localization and mapping (SLAM) is used to predict the dynamic motion path of a moving platform based on the location coordinates and the precise mapping of the physical environment. SLAM has great potential…

Autonomous NavigationAutonomous VehiclesSelf-Driving CarsSimultaneous Localization and Mapping

S3E: A Multi-Robot Multimodal Dataset for Collaborative SLAM

2022-10-25 · Dapeng Feng, Yuhua Qi, Shipeng Zhong, Zhiqiang Chen 외

The burgeoning demand for collaborative robotic systems to execute complex tasks collectively has intensified the research community's focus on advancing simultaneous localization and mapping (SLAM) in a cooperative cont…

DiversitySimultaneous Localization and Mapping