paper-with-me

홈 › Papers

X-3D: Explicit 3D Structure Modeling for Point Cloud Recognition

2024-04-23 · CVPR 2024 1 · Shuofeng Sun, Yongming Rao, Jiwen Lu, Haibin Yan

Numerous prior studies predominantly emphasize constructing relation vectors for individual neighborhood points and generating dynamic kernels for each vector and embedding these into high-dimensional spaces to capture implicit local structures. However, we contend that such implicit high-dimensional structure modeling approch inadequately represents the local geometric structure of point clouds due to the absence of explicit structural information. Hence, we introduce X-3D, an explicit 3D structure modeling approach. X-3D functions by capturing the explicit local structural information within the input 3D space and employing it to produce dynamic kernels with shared weights for all neighborhood points within the current local region. This modeling approach introduces effective geometric prior and significantly diminishes the disparity between the local structure of the embedding space and the original input point cloud, thereby improving the extraction of local features. Experiments show that our method can be used on a variety of methods and achieves state-of-the-art performance on segmentation, classification, detection tasks with lower extra computational cost, such as \textbf{90.7\%} on ScanObjectNN for classification, \textbf{79.2\%} on S3DIS 6 fold and \textbf{74.3\%} on S3DIS Area 5 for segmentation, \textbf{76.3\%} on ScanNetV2 for segmentation and \textbf{64.5\%} mAP , \textbf{46.9\%} mAP on SUN RGB-D and \textbf{69.0\%} mAP , \textbf{51.1\%} mAP on ScanNetV2 . Our code is available at \href{https://github.com/sunshuofeng/X-3D}{https://github.com/sunshuofeng/X-3D}.

📄 PDF Abstract BibTeX arXiv:2404.15010

Code (1)

sunshuofeng/X-3D 공식 구현 pytorch

Tasks

Segmentation

Similar Papers 제목 키워드 기반

Point 4D Transformer Networks for Spatio-Temporal Modeling in Point Cloud Videos

2021-06-19 · CVPR 2021 1 · Hehe Fan, Yi Yang, Mohan Kankanhalli

Point cloud videos exhibit irregularities and lack of order along the spatial dimension where points emerge inconsistently across different frames. To capture the dynamics in point cloud videos, point tracking is usu…

3D Action RecognitionAction RecognitionPoint TrackingSemantic Segmentation

Real-time 3D human action recognition based on Hyperpoint sequence

2021-11-16 · Xing Li, Qian Huang, Zhijian Wang, Zhenjie Hou 외

Real-time 3D human action recognition has broad industrial applications, such as surveillance, human-computer interaction, and healthcare monitoring. By relying on complex spatio-temporal local encoding, most existing po…

3D Action RecognitionAction RecognitionTemporal Action Localization

Flow-based GAN for 3D Point Cloud Generation from a Single Image

2022-10-08 · Yao Wei, George Vosselman, Michael Ying Yang

Generating a 3D point cloud from a single 2D image is of great importance for 3D scene understanding applications. To reconstruct the whole 3D shape of the object shown in the image, the existing deep learning based appr…

Point Cloud GenerationScene Understanding

SRENet: Spectral Re-Entry Network for Point Cloud Action Recognition

2026-06-02 · Qiuxia Wu, Jiarui Lan, Wenxiong Kang, Zhiyong Wang 외 arxiv

Recognizing human actions from point cloud sequences is critical for 3D perception driven applications such as autonomous driving and human-computer interaction. However, the irregular structure and temporal inconsistenc…

Representation LearningAction UnderstandingAction RecognitionAutonomous Driving

Spatiotemporal Learning of Dynamic Gestures from 3D Point Cloud Data

2018-04-24 · Joshua Owoyemi, Koichi Hashimoto

In this paper, we demonstrate an end-to-end spatiotemporal gesture learning approach for 3D point cloud data using a new gestures dataset of point clouds acquired from a 3D sensor. Nine classes of gestures were learned f…

Data AugmentationScene Understanding