paper-with-me

Papers

Complementary Pseudo Multimodal Feature for Point Cloud Anomaly Detection

2023-03-23 · Yunkang Cao, Xiaohao Xu, Weiming Shen

Point cloud (PCD) anomaly detection steadily emerges as a promising research area. This study aims to improve PCD anomaly detection performance by combining handcrafted PCD descriptions with powerful pre-trained 2D neural networks. To this end, this study proposes Complementary Pseudo Multimodal Feature (CPMF) that incorporates local geometrical information in 3D modality using handcrafted PCD descriptors and global semantic information in the generated pseudo 2D modality using pre-trained 2D neural networks. For global semantics extraction, CPMF projects the origin PCD into a pseudo 2D modality containing multi-view images. These images are delivered to pre-trained 2D neural networks for informative 2D modality feature extraction. The 3D and 2D modality features are aggregated to obtain the CPMF for PCD anomaly detection. Extensive experiments demonstrate the complementary capacity between 2D and 3D modality features and the effectiveness of CPMF, with 95.15% image-level AU-ROC and 92.93% pixel-level PRO on the MVTec3D benchmark. Code is available on https://github.com/caoyunkang/CPMF.

📄 PDF Abstract BibTeX arXiv:2303.13194

Code (2)

caoyunkang/CPMF 공식 구현 pytorch
m-3lab/real3d-ad pytorch

Tasks

3D Anomaly Detection and SegmentationAnomaly DetectionDepth Anomaly Detection and Segmentation

Similar Papers 제목 키워드 기반

Real-IAD D3: A Real-World 2D/Pseudo-3D/3D Dataset for Industrial Anomaly Detection

2025-04-19 · CVPR 2025 1 · Wenbing Zhu, Lidong Wang, Ziqing Zhou, Chengjie Wang 외

The increasing complexity of industrial anomaly detection (IAD) has positioned multimodal detection methods as a focal area of machine vision research. However, dedicated multimodal datasets specifically tailored for IAD…

Anomaly Detection

ObitoNet: Multimodal High-Resolution Point Cloud Reconstruction

2024-12-25 · Apoorv Thapliyal, Vinay Lanka, Swathi Baskaran

ObitoNet employs a Cross Attention mechanism to integrate multimodal inputs, where Vision Transformers (ViT) extract semantic features from images and a point cloud tokenizer processes geometric information using Farthes…

DecoderPoint Cloud GenerationPoint cloud reconstruction

LDRFusion: A LiDAR-Dominant multimodal refinement framework for 3D object detection

2025-07-22 · Jijun Wang, Yan Wu, Yujian Mo, Junqiao Zhao 외 arxiv

Existing LiDAR-Camera fusion methods have achieved strong results in 3D object detection. To address the sparsity of point clouds, previous approaches typically construct spatial pseudo point clouds via depth completion …

3D Object DetectionDepth CompletionPoint Clouds

PTA-Det: Point Transformer Associating Point cloud and Image for 3D Object Detection

2023-01-18 · Rui Wan, Tianyun Zhao, Wei Zhao

In autonomous driving, 3D object detection based on multi-modal data has become an indispensable approach when facing complex environments around the vehicle. During multi-modal detection, LiDAR and camera are simultaneo…

3D Object DetectionAutonomous Drivingobject-detectionObject Detection+1

A denoised Mean Teacher for domain adaptive point cloud registration

2023-06-26 · Alexander Bigalke, Mattias P. Heinrich

Point cloud-based medical registration promises increased computational efficiency, robustness to intensity shifts, and anonymity preservation but is limited by the inefficacy of unsupervised learning with similarity met…

Computational EfficiencyDenoisingDomain AdaptationPoint Cloud Registration