paper-with-me

Papers

MANet: Multimodal Attention Network based Point- View fusion for 3D Shape Recognition

2020-02-28 · Yaxin Zhao, Jichao Jiao, Tangkun Zhang

3D shape recognition has attracted more and more attention as a task of 3D vision research. The proliferation of 3D data encourages various deep learning methods based on 3D data. Now there have been many deep learning models based on point-cloud data or multi-view data alone. However, in the era of big data, integrating data of two different modals to obtain a unified 3D shape descriptor is bound to improve the recognition accuracy. Therefore, this paper proposes a fusion network based on multimodal attention mechanism for 3D shape recognition. Considering the limitations of multi-view data, we introduce a soft attention scheme, which can use the global point-cloud features to filter the multi-view features, and then realize the effective fusion of the two features. More specifically, we obtain the enhanced multi-view features by mining the contribution of each multi-view image to the overall shape recognition, and then fuse the point-cloud features and the enhanced multi-view features to obtain a more discriminative 3D shape descriptor. We have performed relevant experiments on the ModelNet40 dataset, and experimental results verify the effectiveness of our method.

📄 PDF Abstract BibTeX arXiv:2002.12573

Code (0)

등록된 구현이 없습니다.

Tasks

3D Shape Recognition

Similar Papers 제목 키워드 기반

MANet: Fine-Tuning Segment Anything Model for Multimodal Remote Sensing Semantic Segmentation

2024-10-15 · Xianping Ma, Xiaokang Zhang, Man-on Pun, Bo Huang

Multimodal remote sensing data, collected from a variety of sensors, provide a comprehensive and integrated perspective of the Earth's surface. By employing multimodal fusion techniques, semantic segmentation offers more…

General KnowledgeSegmentationSemantic Segmentation

Secure Routing Protocol To Mitigate Attacks By Using Blockchain Technology In Manet

2023-04-09 · Nitesh Ghodichor, Raj Thaneeghavl. V, Dinesh Sahu, Gautam Borkar 외

MANET is a collection of mobile nodes that communicate through wireless networks as they move from one point to another. MANET is an infrastructure-less network with a changeable topology; as a result, it is very suscept…

MMANet: Margin-aware Distillation and Modality-aware Regularization for Incomplete Multimodal Learning

2023-04-17 · CVPR 2023 1 · Shicai Wei, Yang Luo, Chunbo Luo

Multimodal learning has shown great potentials in numerous scenes and attracts increasing interest recently. However, it often encounters the problem of missing modality data and thus suffers severe performance degradati…

Semantic Segmentation

FusionBERT: Multi-View Image-3D Retrieval via Cross-Attention Visual Fusion and Normal-Aware 3D Encoder

2026-04-02 · Wei Li, Yufan Ren, Hanqing Jiang, Jianhui Ding 외 arxiv

We propose FusionBERT, a novel multi-view visual fusion framework for image-3D multimodal retrieval. Existing image-3D representation learning methods predominantly focus on feature alignment of a single object image and…

Representation LearningCross-Modal Retrieval

LCPR: A Multi-Scale Attention-Based LiDAR-Camera Fusion Network for Place Recognition

2023-11-06 · Zijie Zhou, Jingyi Xu, Guangming Xiong, Junyi Ma

Place recognition is one of the most crucial modules for autonomous vehicles to identify places that were previously visited in GPS-invalid environments. Sensor fusion is considered an effective method to overcome the we…

Autonomous VehiclesSensor Fusion