paper-with-me

홈 › Papers

VPFNet: Voxel-Pixel Fusion Network for Multi-class 3D Object Detection

2021-11-01 · Chia-Hung Wang, Hsueh-Wei Chen, Li-Chen Fu

Many LiDAR-based methods for detecting large objects, single-class object detection, or under easy situations were claimed to perform quite well. However, their performances of detecting small objects or under hard situations did not surpass those of the fusion-based ones due to failure to leverage the image semantics. In order to elevate the detection performance in a complicated environment, this paper proposes a deep learning (DL)-embedded fusion-based multi-class 3D object detection network which admits both LiDAR and camera sensor data streams, named Voxel-Pixel Fusion Network (VPFNet). Inside this network, a key novel component is called Voxel-Pixel Fusion (VPF) layer, which takes advantage of the geometric relation of a voxel-pixel pair and fuses the voxel features and the pixel features with proper mechanisms. Moreover, several parameters are particularly designed to guide and enhance the fusion effect after considering the characteristics of a voxel-pixel pair. Finally, the proposed method is evaluated on the KITTI benchmark for multi-class 3D object detection task under multilevel difficulty, and is shown to outperform all state-of-the-art methods in mean average precision (mAP). It is also noteworthy that our approach here ranks the first on the KITTI leaderboard for the challenging pedestrian class.

📄 PDF Abstract BibTeX arXiv:2111.00966

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object DetectionObjectobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Variational Probabilistic Fusion Network for RGB-T Semantic Segmentation

2023-07-17 · Baihong Lin, Zengrong Lin, Yulan Guo, Yulan Zhang 외

RGB-T semantic segmentation has been widely adopted to handle hard scenes with poor lighting conditions by fusing different modality features of RGB and thermal images. Existing methods try to find an optimal fusion feat…

SegmentationSemantic SegmentationThermal Image Segmentation

VPFNet: Improving 3D Object Detection with Virtual Point based LiDAR and Stereo Data Fusion

2021-11-29 · Hanqi Zhu, Jiajun Deng, Yu Zhang, Jianmin Ji 외

It has been well recognized that fusing the complementary information from depth-aware LiDAR point clouds and semantic-rich stereo images would benefit 3D object detection. Nevertheless, it is not trivial to explore the …

3D Object DetectionData AugmentationGPUobject-detection+2

SDVRF: Sparse-to-Dense Voxel Region Fusion for Multi-modal 3D Object Detection

2023-04-17 · Binglu Ren, Jianqin Yin

In the perception task of autonomous driving, multi-modal methods have become a trend due to the complementary characteristics of LiDAR point clouds and image data. However, the performance of multi-modal methods is usua…

3D Object DetectionAutonomous Drivingobject-detectionObject Detection

Dense Voxel Fusion for 3D Object Detection

2022-03-02 · Anas Mahmoud, Jordan S. K. Hu, Steven L. Waslander

Camera and LiDAR sensor modalities provide complementary appearance and geometric information useful for detecting 3D objects for autonomous vehicle applications. However, current end-to-end fusion methods are challengin…

3D Object DetectionObjectobject-detectionObject Detection+1

FusionNet: 3D Object Classification Using Multiple Data Representations

2016-07-19 · Vishakh Hegde, Reza Zadeh

High-quality 3D object recognition is an important component of many vision and robotics systems. We tackle the object recognition problem using two data representations, to achieve leading results on the Princeton Model…

3D Object Classification3D Object RecognitionClassificationGeneral Classification+2