paper-with-me

Papers

MultiCorrupt: A Multi-Modal Robustness Dataset and Benchmark of LiDAR-Camera Fusion for 3D Object Detection

2024-02-18 · Till Beemelmanns, Quan Zhang, Christian Geller, Lutz Eckstein

Multi-modal 3D object detection models for automated driving have demonstrated exceptional performance on computer vision benchmarks like nuScenes. However, their reliance on densely sampled LiDAR point clouds and meticulously calibrated sensor arrays poses challenges for real-world applications. Issues such as sensor misalignment, miscalibration, and disparate sampling frequencies lead to spatial and temporal misalignment in data from LiDAR and cameras. Additionally, the integrity of LiDAR and camera data is often compromised by adverse environmental conditions such as inclement weather, leading to occlusions and noise interference. To address this challenge, we introduce MultiCorrupt, a comprehensive benchmark designed to evaluate the robustness of multi-modal 3D object detectors against ten distinct types of corruptions. We evaluate five state-of-the-art multi-modal detectors on MultiCorrupt and analyze their performance in terms of their resistance ability. Our results show that existing methods exhibit varying degrees of robustness depending on the type of corruption and their fusion strategy. We provide insights into which multi-modal design choices make such models robust against certain perturbations. The dataset generation code and benchmark are open-sourced at https://github.com/ika-rwth-aachen/MultiCorrupt.

📄 PDF Abstract BibTeX arXiv:2402.11677

Code (1)

ika-rwth-aachen/multicorrupt 공식 구현

Tasks

3D Object DetectionDataset Generationobject-detectionObject Detection

Similar Papers 제목 키워드 기반

SB-BEVFusion: Enhancing the Robustness against Sensor Malfunction and Corruptions

2026-05-12 · Markus Essl, Marta Moscati, Mubashir Noman, Muhammad Zaigham Zaheer 외 arxiv

Multimodal sensor fusion has demonstrated remarkable performance improvements over unimodal approaches in 3D object detection for autonomous vehicles. Typically, existing methods transform multimodal data from independen…

Autonomous Vehicles3D Object Detection

Benchmarking Multi-modal Semantic Segmentation under Sensor Failures: Missing and Noisy Modality Robustness

2025-03-24 · Chenfei Liao, Kaiyu Lei, Xu Zheng, Junha Moon 외

Multi-modal semantic segmentation (MMSS) addresses the limitations of single-modality data by integrating complementary information across modalities. Despite notable progress, a significant gap persists between research…

BenchmarkingSemantic Segmentation

Benchmarking Robustness of Multimodal Image-Text Models under Distribution Shift

2022-12-15 · JieLin Qiu, Yi Zhu, Xingjian Shi, Florian Wenzel 외

Multimodal image-text models have shown remarkable performance in the past few years. However, evaluating robustness against distribution shifts is crucial before adopting them in real-world applications. In this work, w…

BenchmarkingImage CaptioningImage GenerationImage-text Retrieval+6

Analyzing Modality Robustness in Multimodal Sentiment Analysis

2022-05-30 · NAACL 2022 7 · Devamanyu Hazarika, Yingting Li, Bo Cheng, Shuai Zhao 외

Building robust multimodal models are crucial for achieving reliable deployment in the wild. Despite its importance, less attention has been paid to identifying and improving the robustness of Multimodal Sentiment Analys…

DiagnosticMultimodal Sentiment AnalysisSentiment Analysis

MER 2023: Multi-label Learning, Modality Robustness, and Semi-Supervised Learning

2023-04-18 · Zheng Lian, Haiyang Sun, Licai Sun, Kang Chen 외

The first Multimodal Emotion Recognition Challenge (MER 2023) was successfully held at ACM Multimedia. The challenge focuses on system robustness and consists of three distinct tracks: (1) MER-MULTI, where participants a…

Emotion RecognitionMulti-Label LearningMultimodal Emotion Recognition