IMFNet: Interpretable Multimodal Fusion for Point Cloud Registration
The existing state-of-the-art point descriptor relies on structure information only, which omit the texture information. However, texture information is crucial for our humans to distinguish a scene part. Moreover, the current learning-based point descriptors are all black boxes which are unclear how the original points contribute to the final descriptor. In this paper, we propose a new multimodal fusion method to generate a point cloud registration descriptor by considering both structure and texture information. Specifically, a novel attention-fusion module is designed to extract the weighted texture information for the descriptor extraction. In addition, we propose an interpretable module to explain the original points in contributing to the final descriptor. We use the descriptor element as the loss to backpropagate to the target layer and consider the gradient as the significance of this point to the final descriptor. This paper moves one step further to explainable deep learning in the registration task. Comprehensive experiments on 3DMatch, 3DLoMatch and KITTI demonstrate that the multimodal fusion descriptor achieves state-of-the-art accuracy and improve the descriptor's distinctiveness. We also demonstrate that our interpretable module in explaining the registration descriptor extraction.
Code (1)
Tasks
Point Cloud RegistrationSimilar Papers 제목 키워드 기반
FusionPainting: Multimodal Fusion with Adaptive Attention for 3D Object Detection
Accurate detection of obstacles in 3D is an essential task for autonomous driving and intelligent transportation. In this work, we propose a general multimodal fusion framework FusionPainting to fuse the 2D RGB image and…
3D Object DetectionAutonomous Drivingobject-detectionObject Detection+2DMF-Net: Image-Guided Point Cloud Completion with Dual-Channel Modality Fusion and Shape-Aware Upsampling Transformer
In this paper we study the task of a single-view image-guided point cloud completion. Existing methods have got promising results by fusing the information of image into point cloud explicitly or implicitly. However, giv…
Point Cloud CompletionMultimodal Industrial Anomaly Detection via Hybrid Fusion
2D-based Industrial Anomaly Detection has been widely discussed, however, multimodal industrial anomaly detection based on 3D point clouds and RGB images still has many untouched fields. Existing multimodal industrial an…
3D Anomaly DetectionAnomaly DetectionContrastive LearningRGB+3D Anomaly Detection and SegmentationTriFusion-AE: Language-Guided Depth and LiDAR Fusion for Robust Point Cloud Processing
LiDAR-based perception is central to autonomous driving and robotics, yet raw point clouds remain highly vulnerable to noise, occlusion, and adversarial corruptions. Autoencoders offer a natural framework for denoising a…
Representation LearningAutonomous DrivingPoint CloudsAdaptive and Azimuth-Aware Fusion Network of Multimodal Local Features for 3D Object Detection
This paper focuses on the construction of stronger local features and the effective fusion of image and LiDAR data. We adopt different modalities of LiDAR data to generate richer features and present an adaptive and azim…
3D Object Detectionobject-detectionObject DetectionRegion Proposal