Common Variable Learning and Invariant Representation Learning using Siamese Neural Networks
We consider the statistical problem of learning common source of variability in data which are synchronously captured by multiple sensors, and demonstrate that Siamese neural networks can be naturally applied to this problem. This approach is useful in particular in exploratory, data-driven applications, where neither a model nor label information is available. In recent years, many researchers have successfully applied Siamese neural networks to obtain an embedding of data which corresponds to a "semantic similarity". We present an interpretation of this "semantic similarity" as learning of equivalence classes. We discuss properties of the embedding obtained by Siamese networks and provide empirical results that demonstrate the ability of Siamese networks to learn common variability.
Code (0)
등록된 구현이 없습니다.
Tasks
Representation LearningSemantic SimilaritySemantic Textual SimilaritySimilar Papers 제목 키워드 기반
Single microphone speaker extraction using unified time-frequency Siamese-Unet
In this paper we present a unified time-frequency method for speaker extraction in clean and noisy conditions. Given a mixed signal, along with a reference signal, the common approaches for extracting the desired speaker…
blind source separationDecoderDeep Intra-Image Contrastive Learning for Weakly Supervised One-Step Person Search
Weakly supervised person search aims to perform joint pedestrian detection and re-identification (re-id) with only person bounding-box annotations. Recently, the idea of contrastive learning is initially applied to weakl…
Contrastive LearningPedestrian DetectionPerson SearchLearning an MR acquisition-invariant representation using Siamese neural networks
Generalization of voxelwise classifiers is hampered by differences between MRI-scanners, e.g. different acquisition protocols and field strengths. To address this limitation, we propose a Siamese neural network (MRAI-NET…
Siamese DETR
Recent self-supervised methods are mainly designed for representation learning with the base model, e.g., ResNets or ViTs. They cannot be easily transferred to DETR, with task-specific Transformer modules. In this work, …
MULTI-VIEW LEARNINGRepresentation LearningSampling strategies in Siamese Networks for unsupervised speech representation learning
Recent studies have investigated siamese network architectures for learning invariant speech representations using same-different side information at the word level. Here we investigate systematically an often ignored co…
Representation LearningSpeech Representation Learning