Shared Manifold Learning Using a Triplet Network for Multiple Sensor Translation and Fusion with Missing Data
Heterogeneous data fusion can enhance the robustness and accuracy of an algorithm on a given task. However, due to the difference in various modalities, aligning the sensors and embedding their information into discriminative and compact representations is challenging. In this paper, we propose a Contrastive learning based MultiModal Alignment Network (CoMMANet) to align data from different sensors into a shared and discriminative manifold where class information is preserved. The proposed architecture uses a multimodal triplet autoencoder to cluster the latent space in such a way that samples of the same classes from each heterogeneous modality are mapped close to each other. Since all the modalities exist in a shared manifold, a unified classification framework is proposed. The resulting latent space representations are fused to perform more robust and accurate classification. In a missing sensor scenario, the latent space of one sensor is easily and efficiently predicted using another sensor's latent space, thereby allowing sensor translation. We conducted extensive experiments on a manually labeled multimodal dataset containing hyperspectral data from AVIRIS-NG and NEON, and LiDAR (light detection and ranging) data from NEON. Lastly, the model is validated on two benchmark datasets: Berlin Dataset (hyperspectral and synthetic aperture radar) and MUUFL Gulfport Dataset (hyperspectral and LiDAR). A comparison made with other methods demonstrates the superiority of this method. We achieved a mean overall accuracy of 94.3% on the MUUFL dataset and the best overall accuracy of 71.26% on the Berlin dataset, which is better than other state-of-the-art approaches.
Code (0)
등록된 구현이 없습니다.
Tasks
Contrastive LearningTripletMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Variational learning across domains with triplet information
The work investigates deep generative models, which allow us to use training data from one domain to build a model for another domain. We propose the Variational Bi-domain Triplet Autoencoder (VBTA) that learns a joint d…
Cross-Lingual Document ClassificationDocument ClassificationGeneral ClassificationImage Generation+4Variational learning across domains with triplet information
The work investigates deep generative models, which allow us to use training data from one domain to build a model for another domain. We propose the Variational Bi-domain Triplet Autoencoder (VBTA) that learns a joint d…
Cross-Lingual Document ClassificationDocument ClassificationImage GenerationImage-to-Image Translation+3Jointly Extracting Multiple Triplets with Multilayer Translation Constraints
Triplets extraction is an essential and pivotal step in automatic knowledge base construction, which captures structural information from unstructured text corpus. Conventional extraction models use a pipeline of named e…
Knowledge Base Constructionnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+4Self-Supervised Anomaly Detection of Rogue Soil Moisture Sensors
IoT data is a central element in the successful digital transformation of agriculture. However, IoT data comes with its own set of challenges. E.g., the risk of data contamination due to rogue sensors. A sensor is consid…
Anomaly DetectionDynamic Time WarpingSelf-Supervised Anomaly DetectionSupervised Anomaly Detection+1Generative Imagination Elevates Machine Translation
There are common semantics shared across text and images. Given a sentence in a source language, whether depicting the visual scene helps translation into a target language? Existing multimodal neural machine translation…
Machine TranslationMultimodal Machine TranslationSentenceTransfer Learning+1