paper-with-me

홈 › Papers

ViewFormer: View Set Attention for Multi-view 3D Shape Understanding

2023-04-29 · Hongyu Sun, Yongcai Wang, Peng Wang, Xudong Cai, Deying Li

This paper presents ViewFormer, a simple yet effective model for multi-view 3d shape recognition and retrieval. We systematically investigate the existing methods for aggregating multi-view information and propose a novel ``view set" perspective, which minimizes the relation assumption about the views and releases the representation flexibility. We devise an adaptive attention model to capture pairwise and higher-order correlations of the elements in the view set. The learned multi-view correlations are aggregated into an expressive view set descriptor for recognition and retrieval. Experiments show the proposed method unleashes surprising capabilities across different tasks and datasets. For instance, with only 2 attention blocks and 4.8M learnable parameters, ViewFormer reaches 98.8% recognition accuracy on ModelNet40 for the first time, exceeding previous best method by 1.1% . On the challenging RGBD dataset, our method achieves 98.4% recognition accuracy, which is a 4.1% absolute improvement over the strongest baseline. ViewFormer also sets new records in several evaluation dimensions of 3D shape retrieval defined on the SHREC'17 benchmark.

📄 PDF Abstract BibTeX arXiv:2305.00161

Code (0)

등록된 구현이 없습니다.

Tasks

3D Shape Recognition3D Shape RetrievalRetrieval

Similar Papers 제목 키워드 기반

ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers

2024-05-07 · Jinke Li, Xiao He, Chonghua Zhou, Xiaoqiang Cheng 외

3D occupancy, an advanced perception technology for driving scenarios, represents the entire scene without distinguishing between foreground and background by quantifying the physical space into a grid map. The widely ad…

3D Object Detectionobject-detectionObject Detection

ViewFormer: NeRF-free Neural Rendering from Few Images Using Transformers

2022-03-18 · Jonáš Kulhánek, Erik Derner, Torsten Sattler, Robert Babuška

Novel view synthesis is a long-standing problem. In this work, we consider a variant of the problem where we are given only a few context views sparsely covering a scene or an object. The goal is to predict novel viewpoi…

Camera Pose EstimationNeRFNeural RenderingNovel View Synthesis+1

Fine-Grained 3D Shape Classification with Hierarchical Part-View Attentions

2020-05-26 · Xinhai Liu, Zhizhong Han, Yu-Shen Liu, Matthias Zwicker

Fine-grained 3D shape classification is important for shape understanding and analysis, which poses a challenging research problem. However, the studies on the fine-grained 3D shape classification have rarely been explor…

3D Shape ClassificationGeneral ClassificationRegion ProposalSemantic Part Detection

MANet: Multimodal Attention Network based Point- View fusion for 3D Shape Recognition

2020-02-28 · Yaxin Zhao, Jichao Jiao, Tangkun Zhang

3D shape recognition has attracted more and more attention as a task of 3D vision research. The proliferation of 3D data encourages various deep learning methods based on 3D data. Now there have been many deep learning m…

3D Shape Recognition

Learning Attentive and Hierarchical Representations for 3D Shape Recognition

2020-08-01 · ECCV 2020 8 · Jiaxin Chen, Jie Qin, Yuming Shen, Li Liu 외

This paper proposes a novel method for 3D shape representation learning, namely Hyperbolic Embedded Attentive Representation (HEAR). Different from existing multi-view based methods, HEAR develops a unified framework to …

3D Shape Classification3D Shape Recognition3D Shape Representation3D Shape Retrieval+2