paper-with-me

Papers

DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection

2025-11-13 · Feiyang Jia, Caiyan Jia, Ailin Liu, Shaoqing Xu, Qiming Xia, Lin Liu, Lei Yang, Yan Gong, Ziying Song arxiv

As a critical task in autonomous driving perception systems, 3D object detection is used to identify and track key objects, such as vehicles and pedestrians. However, detecting distant, small, or occluded objects (hard instances) remains a challenge, which directly compromises the safety of autonomous driving systems. We observe that existing multi-modal 3D object detection methods often follow a single-guided paradigm, failing to account for the differences in information density of hard instances between modalities. In this work, we propose DGFusion, based on the Dual-guided paradigm, which fully inherits the advantages of the Point-guide-Image paradigm and integrates the Image-guide-Point paradigm to address the limitations of the single paradigms. The core of DGFusion, the Difficulty-aware Instance Pair Matcher (DIPM), performs instance-level feature matching based on difficulty to generate easy and hard instance pairs, while the Dual-guided Modules exploit the advantages of both pair types to enable effective multi-modal feature fusion. Experimental results demonstrate that our DGFusion outperforms the baseline methods, with respective improvements of +1.0\% mAP, +0.8\% NDS, and +1.3\% average recall on nuScenes. Extensive experiments demonstrate consistent robustness gains for hard instance detection across ego-distance, size, visibility, and small-scale training scenarios.

📄 PDF Abstract BibTeX arXiv:2511.10035

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object DetectionAutonomous Driving

Results from the Paper

RankTaskDatasetModelMetrics
#3 3D Object Detection nuScenes DGFusion NDS: 0.8

Similar Papers 제목 키워드 기반

DGFusion: Depth-Guided Sensor Fusion for Robust Semantic Perception

2025-09-11 · Tim Broedermannn, Christos Sakaridis, Luigi Piccinelli, Wim Abbeloos 외 arxiv

Robust semantic perception for autonomous vehicles relies on effectively combining multiple sensors with complementary strengths and weaknesses. State-of-the-art sensor fusion approaches to semantic perception often trea…

Semantic SegmentationAutonomous Vehicles

Dual-Domain Perspective on Degradation-Aware Fusion: A VLM-Guided Robust Infrared and Visible Image Fusion Framework

2025-09-05 · Tianpei Zhang, Jufeng Zhao, Yiming Zhu, Guangmang Cui arxiv

Most existing infrared-visible image fusion (IVIF) methods assume high-quality inputs, and therefore struggle to handle dual-source degraded scenarios, typically requiring manual selection and sequential application of m…

SeaDATE: Remedy Dual-Attention Transformer with Semantic Alignment via Contrast Learning for Multimodal Object Detection

2024-10-15 · Shuhan Dong, Yunsong Li, Weiying Xie, Jiaqing Zhang 외

Multimodal object detection leverages diverse modal information to enhance the accuracy and robustness of detectors. By learning long-term dependencies, Transformer can effectively integrate multimodal features in the fe…

Contrastive Learningobject-detectionObject Detection

A Tri-attention Fusion Guided Multi-modal Segmentation Network

2021-11-02 · Tongxue Zhou, Su Ruan, Pierre Vera, Stéphane Canu

In the field of multimodal segmentation, the correlation between different modalities can be considered for improving the segmentation results. Considering the correlation between different MR modalities, in this paper, …

Brain Tumor SegmentationSegmentationTumor Segmentation

DMF-Net: Image-Guided Point Cloud Completion with Dual-Channel Modality Fusion and Shape-Aware Upsampling Transformer

2024-06-25 · Aihua Mao, Yuxuan Tang, Jiangtao Huang, Ying He

In this paper we study the task of a single-view image-guided point cloud completion. Existing methods have got promising results by fusing the information of image into point cloud explicitly or implicitly. However, giv…

Point Cloud Completion