Choosing Smartly: Adaptive Multimodal Fusion for Object Detection in Changing Environments
Object detection is an essential task for autonomous robots operating in dynamic and changing environments. A robot should be able to detect objects in the presence of sensor noise that can be induced by changing lighting conditions for cameras and false depth readings for range sensors, especially RGB-D cameras. To tackle these challenges, we propose a novel adaptive fusion approach for object detection that learns weighting the predictions of different sensor modalities in an online manner. Our approach is based on a mixture of convolutional neural network (CNN) experts and incorporates multiple modalities including appearance, depth and motion. We test our method in extensive robot experiments, in which we detect people in a combined indoor and outdoor scenario from RGB-D data, and we demonstrate that our method can adapt to harsh lighting changes and severe camera motion blur. Furthermore, we present a new RGB-D dataset for people detection in mixed in- and outdoor environments, recorded with a mobile robot. Code, pretrained models and dataset are available at http://adaptivefusion.cs.uni-freiburg.de
Code (1)
Tasks
object-detectionObject DetectionSimilar Papers 제목 키워드 기반
FGU3R: Fine-Grained Fusion via Unified 3D Representation for Multimodal 3D Object Detection
Multimodal 3D object detection has garnered considerable interest in autonomous driving. However, multimodal detectors suffer from dimension mismatches that derive from fusing 3D points with 2D pixels coarsely, which lea…
3D Object DetectionAutonomous Drivingmultimodal interactionobject-detection+1MixMAS: A Framework for Sampling-Based Mixer Architecture Search for Multimodal Fusion and Learning
Choosing a suitable deep learning architecture for multimodal data fusion is a challenging task, as it requires the effective integration and processing of diverse data types, each with distinct structures and characteri…
BenchmarkingWeakly Misalignment-free Adaptive Feature Alignment for UAVs-based Multimodal Object Detection
Visible-infrared (RGB-IR) image fusion has shown great potentials in object detection based on unmanned aerial vehicles (UAVs). However the weakly misalignment problem between multimodal image pairs limits its perfor…
2D Object DetectionObjectobject-detectionObject DetectionFusionPainting: Multimodal Fusion with Adaptive Attention for 3D Object Detection
Accurate detection of obstacles in 3D is an essential task for autonomous driving and intelligent transportation. In this work, we propose a general multimodal fusion framework FusionPainting to fuse the 2D RGB image and…
3D Object DetectionAutonomous Drivingobject-detectionObject Detection+2Adaptive and Azimuth-Aware Fusion Network of Multimodal Local Features for 3D Object Detection
This paper focuses on the construction of stronger local features and the effective fusion of image and LiDAR data. We adopt different modalities of LiDAR data to generate richer features and present an adaptive and azim…
3D Object Detectionobject-detectionObject DetectionRegion Proposal