paper-with-me

홈 › Papers

PV-SSD: A Multi-Modal Point Cloud Feature Fusion Method for Projection Features and Variable Receptive Field Voxel Features

2023-08-13 · Yongxin Shao, Aihong Tan, Zhetao Sun, Enhui Zheng, Tianhong Yan, Peng Liao

LiDAR-based 3D object detection and classification is crucial for autonomous driving. However, real-time inference from extremely sparse 3D data is a formidable challenge. To address this problem, a typical class of approaches transforms the point cloud cast into a regular data representation (voxels or projection maps). Then, it performs feature extraction with convolutional neural networks. However, such methods often result in a certain degree of information loss due to down-sampling or over-compression of feature information. This paper proposes a multi-modal point cloud feature fusion method for projection features and variable receptive field voxel features (PV-SSD) based on projection and variable voxelization to solve the information loss problem. We design a two-branch feature extraction structure with a 2D convolutional neural network to extract the point cloud's projection features in bird's-eye view to focus on the correlation between local features. A voxel feature extraction branch is used to extract local fine-grained features. Meanwhile, we propose a voxel feature extraction method with variable sensory fields to reduce the information loss of voxel branches due to downsampling. It avoids missing critical point information by selecting more useful feature points based on feature point weights for the detection task. In addition, we propose a multi-modal feature fusion module for point clouds. To validate the effectiveness of our method, we tested it on the KITTI dataset and ONCE dataset.

📄 PDF Abstract BibTeX arXiv:2308.06791

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object DetectionAutonomous Drivingobject-detectionObject Detection

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

ProtoTransfer: Cross-Modal Prototype Transfer for Point Cloud Segmentation

2023-01-01 · ICCV 2023 1 · Pin Tang, Hai-Ming Xu, Chao Ma

Knowledge transfer from multi-modal, i.e., LiDAR points and images, to a single LiDAR modal can take advantage of complimentary information from modal-fusion but keep a single modal inference speed, showing a promisi…

Autonomous DrivingPoint Cloud SegmentationSemantic SegmentationTransfer Learning

Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection

2022-10-18 · Xin Li, Botian Shi, Yuenan Hou, Xingjiao Wu 외

Multi-modal 3D object detection has been an active research topic in autonomous driving. Nevertheless, it is non-trivial to explore the cross-modal feature fusion between sparse 3D points and dense 2D pixels. Recent appr…

3D Object DetectionAutonomous Drivingobject-detectionObject Detection

MMF-Track: Multi-modal Multi-level Fusion for 3D Single Object Tracking

2023-05-11 · Zhiheng Li, Yubo Cui, Yu Lin, Zheng Fang

3D single object tracking plays a crucial role in computer vision. Mainstream methods mainly rely on point clouds to achieve geometry matching between target template and search area. However, textureless and incomplete …

3D Single Object TrackingObject Tracking

Interactive Multi-scale Fusion of 2D and 3D Features for Multi-object Tracking

2022-03-30 · Guangming Wang, Chensheng Peng, Jinpeng Zhang, Hesheng Wang

Multiple object tracking (MOT) is a significant task in achieving autonomous driving. Traditional works attempt to complete this task, either based on point clouds (PC) collected by LiDAR, or based on images captured fro…

Autonomous DrivingMulti-Object TrackingMultiple Object TrackingObject Tracking

FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators

2023-10-05 · Haiping Wang, YuAn Liu, Bing Wang, Yujing Sun 외

Matching cross-modality features between images and point clouds is a fundamental problem for image-to-point cloud registration. However, due to the modality difference between images and points, it is difficult to learn…

Image to Point Cloud RegistrationMetric LearningPoint Cloud Registration