paper-with-me

Papers

Embedded Deep Bilinear Interactive Information and Selective Fusion for Multi-view Learning

2020-07-13 · Jinglin Xu, Wenbin Li, Jiantao Shen, Xinwang Liu, Peicheng Zhou, Xiangsen Zhang, Xiwen Yao, Junwei Han

As a concrete application of multi-view learning, multi-view classification improves the traditional classification methods significantly by integrating various views optimally. Although most of the previous efforts have been demonstrated the superiority of multi-view learning, it can be further improved by comprehensively embedding more powerful cross-view interactive information and a more reliable multi-view fusion strategy in intensive studies. To fulfill this goal, we propose a novel multi-view learning framework to make the multi-view classification better aimed at the above-mentioned two aspects. That is, we seamlessly embed various intra-view information, cross-view multi-dimension bilinear interactive information, and a new view ensemble mechanism into a unified framework to make a decision via the optimization. In particular, we train different deep neural networks to learn various intra-view representations, and then dynamically learn multi-dimension bilinear interactive information from different bilinear similarities via the bilinear function between views. After that, we adaptively fuse the representations of multiple views by flexibly tuning the parameters of the view-weight, which not only avoids the trivial solution of weight but also provides a new way to select a few discriminative views that are beneficial to make a decision for the multi-view classification. Extensive experiments on six publicly available datasets demonstrate the effectiveness of the proposed method.

📄 PDF Abstract BibTeX arXiv:2007.06143

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral ClassificationMULTI-VIEW LEARNING

Similar Papers 제목 키워드 기반

Deep Fusion: An Attention Guided Factorized Bilinear Pooling for Audio-video Emotion Recognition

2019-01-15 · Yuanyuan Zhang, Zi-Rui Wang, Jun Du

Automatic emotion recognition (AER) is a challenging task due to the abstract concept and multiple expressions of emotion. Although there is no consensus on a definition, human emotional states usually can be apperceived…

Emotion RecognitionVideo Emotion Recognition

IMF: Interactive Multimodal Fusion Model for Link Prediction

2023-03-20 · Xinhang Li, Xiangyu Zhao, Jiaxing Xu, Yong Zhang 외

Link prediction aims to identify potential missing triples in knowledge graphs. To get better results, some recent studies have introduced multimodal information to link prediction. However, these methods utilize multimo…

Contrastive LearningKnowledge GraphsLink PredictionPrediction

Bilinear Attention Networks

2018-05-21 · NeurIPS 2018 12 · Jin-Hwa Kim, Jaehyun Jun, Byoung-Tak Zhang

Attention networks in multimodal learning provide an efficient way to utilize given visual information selectively. However, the computational cost to learn attention distributions for every pair of multimodal input chan…

Visual Question AnsweringVisual Question Answering (VQA)

ZK-WAGON: Imperceptible Watermark for Image Generation Models using ZK-SNARKs

2025-10-02 · Aadarsh Anantha Ramakrishnan, Shubham Agarwal, Selvanayagam S, Kunwar Singh arxiv

As image generation models grow increasingly powerful and accessible, concerns around authenticity, ownership, and misuse of synthetic media have become critical. The ability to generate lifelike images indistinguishable…

Image Generation

Towards Joint Intent Detection and Slot Filling via Higher-order Attention

2021-09-18 · Dongsheng Chen, Zhiqi Huang, Xian Wu, Shen Ge 외

Intent detection (ID) and Slot filling (SF) are two major tasks in spoken language understanding (SLU). Recently, attention mechanism has been shown to be effective in jointly optimizing these two tasks in an interactive…

Intent Detectionslot-fillingSlot FillingSpoken Language Understanding