paper-with-me

Papers

TS-Net: Combining modality specific and common features for multimodal patch matching

2018-06-05 · Sovann En, Alexis Lechervy, Frédéric Jurie

Multimodal patch matching addresses the problem of finding the correspondences between image patches from two different modalities, e.g. RGB vs sketch or RGB vs near-infrared. The comparison of patches of different modalities can be done by discovering the information common to both modalities (Siamese like approaches) or the modality-specific information (Pseudo-Siamese like approaches). We observed that none of these two scenarios is optimal. This motivates us to propose a three-stream architecture, dubbed as TS-Net, combining the benefits of the two. In addition, we show that adding extra constraints in the intermediate layers of such networks further boosts the performance. Experimentations on three multimodal datasets show significant performance gains in comparison with Siamese and Pseudo-Siamese networks.

📄 PDF Abstract BibTeX arXiv:1806.01550

Code (1)

ensv/TS-Net 공식 구현 tf

Tasks

Multimodal Patch MatchingPatch Matching

Similar Papers 제목 키워드 기반

Adversarial Multimodal Representation Learning for Click-Through Rate Prediction

2020-03-07 · Xiang Li, Chao Wang, Jiwei Tan, Xiaoyi Zeng 외

For better user experience and business effectiveness, Click-Through Rate (CTR) prediction has been one of the most important tasks in E-commerce. Although extensive CTR prediction models have been proposed, learning goo…

Click-Through Rate PredictionPredictionRepresentation Learning

Progressive Modality Reinforcement for Human Multimodal Emotion Recognition From Unaligned Multimodal Sequences

2021-06-19 · CVPR 2021 1 · Fengmao Lv, Xiang Chen, Yanyong Huang, Lixin Duan 외

Human multimodal emotion recognition involves time-series data of different modalities, such as natural language, visual motions, and acoustic behaviors. Due to the variable sampling rates for sequences from differen…

Emotion RecognitionMultimodal Emotion RecognitionTime SeriesTime Series Analysis

CentralNet: a Multilayer Approach for Multimodal Fusion

2018-08-22 · Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux, Frédéric Jurie

This paper proposes a novel multimodal fusion approach, aiming to produce best possible decisions by integrating information coming from multiple media. While most of the past multimodal approaches either work by project…

Multi-Task Learning

Multi-Modality Collaborative Learning for Sentiment Analysis

2025-01-21 · Shanmin Wang, Chengguang Liu, Qingshan Liu

Multimodal sentiment analysis (MSA) identifies individuals' sentiment states in videos by integrating visual, audio, and text modalities. Despite progress in existing methods, the inherent modality heterogeneity limits t…

Multimodal Sentiment AnalysisSentiment Analysis

Prior-guided Fusion of Multimodal Features for Change Detection from Optical-SAR Images

2026-04-07 · Xuanguang Liu, Lei Ding, Yujie Li, Chenguang Dai 외 arxiv

Multimodal change detection (MMCD) identifies changed areas in multimodal remote sensing data, demonstrating significant application value in land use monitoring and urban sustainable development. However, literature MMC…

Feature ImportanceChange Detection