paper-with-me

Papers

CentralNet: a Multilayer Approach for Multimodal Fusion

2018-08-22 · Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux, Frédéric Jurie

This paper proposes a novel multimodal fusion approach, aiming to produce best possible decisions by integrating information coming from multiple media. While most of the past multimodal approaches either work by projecting the features of different modalities into the same space, or by coordinating the representations of each modality through the use of constraints, our approach borrows from both visions. More specifically, assuming each modality can be processed by a separated deep convolutional network, allowing to take decisions independently from each modality, we introduce a central network linking the modality specific networks. This central network not only provides a common feature embedding but also regularizes the modality specific networks through the use of multi-task learning. The proposed approach is validated on 4 different computer vision tasks on which it consistently improves the accuracy of existing multimodal fusion approaches.

📄 PDF Abstract BibTeX arXiv:1808.07275

Code (2)

jhaprince/multibully pytorch
mengmenm/SMIL pytorch

Tasks

Multi-Task Learning

Similar Papers 제목 키워드 기반

Multi-Level Sensor Fusion with Deep Learning

2018-11-05 · Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux, Frédéric Jurie

In the context of deep learning, this article presents an original deep network, namely CentralNet, for the fusion of information coming from different sensors. This approach is designed to efficiently and automatically …

Deep LearningSensor Fusion

MultiFusionNet: Multilayer Multimodal Fusion of Deep Neural Networks for Chest X-Ray Image Classification

2024-01-01 · Saurabh Agarwal, K. V. Arya, Yogesh Kumar Meena

Chest X-ray imaging is a critical diagnostic tool for identifying pulmonary diseases. However, manual interpretation of these images is time-consuming and error-prone. Automated systems utilizing convolutional neural net…

Diagnosticimage-classificationImage Classification

A Graph Framework for Multimodal Medical Information Processing

2017-02-22 · Drakopoulos Georgios, Megalooikonomou Vasileios

Multimodal medical information processing is currently the epicenter of intense interdisciplinary research, as proper data fusion may lead to more accurate diagnoses. Moreover, multimodality may disambiguate cases of co-…

Graph Ranking

REMOTE: A Unified Multimodal Relation Extraction Framework with Multilevel Optimal Transport and Mixture-of-Experts

2025-09-05 · Xinkui Lin, Yongxiu Xu, Minghao Tang, Shilong Zhang 외 arxiv

Multimodal relation extraction (MRE) is a crucial task in the fields of Knowledge Graph and Multimedia, playing a pivotal role in multimodal knowledge graph construction. However, existing methods are typically limited t…

Relation Extraction

A Multimodal Learning Framework for Comprehensive 3D Mineral Prospectivity Modeling with Jointly Learned Structure-Fluid Relationships

2023-09-06 · Yang Zheng, Hao Deng, Ruisheng Wang, Jingjie Wu

This study presents a novel multimodal fusion model for three-dimensional mineral prospectivity mapping (3D MPM), effectively integrating structural and fluid information through a deep network architecture. Leveraging C…

Data IntegrationDecision Making