paper-with-me

홈 › Papers

InsFusion: Rethink Instance-level LiDAR-Camera Fusion for 3D Object Detection

2025-09-10 · Zhongyu Xia, Hansong Yang, Yongtao Wang arxiv

Three-dimensional Object Detection from multi-view cameras and LiDAR is a crucial component for autonomous driving and smart transportation. However, in the process of basic feature extraction, perspective transformation, and feature fusion, noise and error will gradually accumulate. To address this issue, we propose InsFusion, which can extract proposals from both raw and fused features and utilizes these proposals to query the raw features, thereby mitigating the impact of accumulated errors. Additionally, by incorporating attention mechanisms applied to the raw features, it thereby mitigates the impact of accumulated errors. Experiments on the nuScenes dataset demonstrate that InsFusion is compatible with various advanced baseline methods and delivers new state-of-the-art performance for 3D object detection.

📄 PDF Abstract BibTeX arXiv:2509.08374

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object DetectionAutonomous Driving

Similar Papers 제목 키워드 기반

SemanticBEVFusion: Rethink LiDAR-Camera Fusion in Unified Bird's-Eye View Representation for 3D Object Detection

2022-12-09 · Qi Jiang, Hao Sun, Xi Zhang

LiDAR and camera are two essential sensors for 3D object detection in autonomous driving. LiDAR provides accurate and reliable 3D geometry information while the camera provides rich texture with color. Despite the increa…

3D geometry3D Object DetectionAutonomous Drivingobject-detection+1

ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection

2024-05-27 · Ziying Song, Feiyang Jia, Hongyu Pan, Yadan Luo 외

In the field of 3D object detection tasks, fusing heterogeneous features from LiDAR and camera sensors into a unified Bird's Eye View (BEV) representation is a widely adopted paradigm. However, existing methods are often…

3D Object DetectionContrastive LearningDepth EstimationGraph Matching+2

VoxDet: Rethinking 3D Semantic Occupancy Prediction as Dense Object Detection

2025-06-05 · Wuyang Li, Zhu Yu, Alexandre Alahi

3D semantic occupancy prediction aims to reconstruct the 3D geometry and semantics of the surrounding environment. With dense voxel labels, prior works typically formulate it as a dense segmentation task, independently c…

3D geometry3D Semantic Occupancy PredictionDense Object Detectionobject-detection+1

RCM-Fusion: Radar-Camera Multi-Level Fusion for 3D Object Detection

2023-07-17 · Jisong Kim, Minjae Seong, Geonho Bang, Dongsuk Kum 외

While LiDAR sensors have been successfully applied to 3D object detection, the affordability of radar and camera sensors has led to a growing interest in fusing radars and cameras for 3D object detection. However, previo…

3D Object DetectionObjectobject-detectionObject Detection

Sparse Beats Dense: Rethinking Supervision in Radar-Camera Depth Completion

2023-12-01 · Huadong Li, Minhao Jing, Jiajun Liang, Haoqiang Fan 외

It is widely believed that sparse supervision is worse than dense supervision in the field of depth completion, but the underlying reasons for this are rarely discussed. To this end, we revisit the task of radar-camera d…

Depth CompletionDepth EstimationDepth PredictionPosition