paper-with-me

홈 › Papers

P2P: Tuning Pre-trained Image Models for Point Cloud Analysis with Point-to-Pixel Prompting

2022-08-04 · Ziyi Wang, Xumin Yu, Yongming Rao, Jie zhou, Jiwen Lu

Nowadays, pre-training big models on large-scale datasets has become a crucial topic in deep learning. The pre-trained models with high representation ability and transferability achieve a great success and dominate many downstream tasks in natural language processing and 2D vision. However, it is non-trivial to promote such a pretraining-tuning paradigm to the 3D vision, given the limited training data that are relatively inconvenient to collect. In this paper, we provide a new perspective of leveraging pre-trained 2D knowledge in 3D domain to tackle this problem, tuning pre-trained image models with the novel Point-to-Pixel prompting for point cloud analysis at a minor parameter cost. Following the principle of prompting engineering, we transform point clouds into colorful images with geometry-preserved projection and geometry-aware coloring to adapt to pre-trained image models, whose weights are kept frozen during the end-to-end optimization of point cloud analysis tasks. We conduct extensive experiments to demonstrate that cooperating with our proposed Point-to-Pixel Prompting, better pre-trained image model will lead to consistently better performance in 3D vision. Enjoying prosperous development from image pre-training field, our method attains 89.3% accuracy on the hardest setting of ScanObjectNN, surpassing conventional point cloud models with much fewer trainable parameters. Our framework also exhibits very competitive performance on ModelNet classification and ShapeNet Part Segmentation. Code is available at https://github.com/wangzy22/P2P.

📄 PDF Abstract BibTeX arXiv:2208.02812

Code (1)

wangzy22/P2P 공식 구현 pytorch

Tasks

3D Part Segmentation3D Point Cloud Classification

Similar Papers 제목 키워드 기반

Adaptive Point-Prompt Tuning: Fine-Tuning Heterogeneous Foundation Models for 3D Point Cloud Analysis

2025-08-30 · Mengke Li, Lihao Chen, Peng Zhang, Yiu-ming Cheung 외 arxiv

Parameter-efficient fine-tuning strategies for foundation models in 1D textual and 2D visual analysis have demonstrated remarkable efficacy. However, due to the scarcity of point cloud data, pre-training large 3D models …

parameter-efficient fine-tuningPoint Clouds

Image2Point: 3D Point-Cloud Understanding with 2D Image Pretrained Models

2021-06-08 · Chenfeng Xu, Shijia Yang, Tomer Galanti, Bichen Wu 외

3D point-clouds and 2D images are different visual representations of the physical world. While human vision can understand both representations, computer vision models designed for 2D image and 3D point-cloud understand…

3D Point Cloud ClassificationPoint Cloud ClassificationScene Segmentation

Dynamic Adapter Meets Prompt Tuning: Parameter-Efficient Transfer Learning for Point Cloud Analysis

2024-03-03 · CVPR 2024 1 · Xin Zhou, Dingkang Liang, Wei Xu, Xingkui Zhu 외

Point cloud analysis has achieved outstanding performance by transferring point cloud pre-trained models. However, existing methods for model adaptation usually update all model parameters, i.e., full fine-tuning paradig…

3D Parameter-Efficient Fine-Tuning for ClassificationGPUTransfer Learning

A Hybrid Generative and Discriminative PointNet on Unordered Point Sets

2024-04-19 · Yang Ye, Shihao Ji

As point cloud provides a natural and flexible representation usable in myriad applications (e.g., robotics and self-driving cars), the ability to synthesize point clouds for analysis becomes crucial. Recently, Xie et al…

image-classificationImage ClassificationPoint Cloud ClassificationPoint Cloud Generation+1

Adapt PointFormer: 3D Point Cloud Analysis via Adapting 2D Visual Transformers

2024-07-18 · Mengke Li, Da Li, Guoqing Yang, Yiu-ming Cheung 외

Pre-trained large-scale models have exhibited remarkable efficacy in computer vision, particularly for 2D image analysis. However, when it comes to 3D point clouds, the constrained accessibility of data, in contrast to t…