paper-with-me

Papers

Learning State Representations in Complex Systems with Multimodal Data

2018-11-27 · Pavel Solovev, Vladimir Aliev, Pavel Ostyakov, Gleb Sterkin, Elizaveta Logacheva, Stepan Troeshestov, Roman Suvorov, Anton Mashikhin, Oleg Khomenko, Sergey I. Nikolenko

Representation learning becomes especially important for complex systems with multimodal data sources such as cameras or sensors. Recent advances in reinforcement learning and optimal control make it possible to design control algorithms on these latent representations, but the field still lacks a large-scale standard dataset for unified comparison. In this work, we present a large-scale dataset and evaluation framework for representation learning for the complex task of landing an airplane. We implement and compare several approaches to representation learning on this dataset in terms of the quality of simple supervised learning tasks and disentanglement scores. The resulting representations can be used for further tasks such as anomaly detection, optimal control, model-based reinforcement learning, and other applications.

📄 PDF Abstract BibTeX arXiv:1811.11067

Code (0)

등록된 구현이 없습니다.

Tasks

Anomaly DetectionDisentanglementModel-based Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Representation Learning

Similar Papers 제목 키워드 기반

Improving Multimodal fusion via Mutual Dependency Maximisation

2021-08-31 · EMNLP 2021 11 · Pierre Colombo, Emile Chapuis, Matthieu Labeau, Chloe Clavel

Multimodal sentiment analysis is a trending area of research, and the multimodal fusion is one of its most active topic. Acknowledging humans communicate through a variety of channels (i.e visual, acoustic, linguistic), …

Multimodal Sentiment AnalysisSentiment Analysis

Do Recommender Systems Really Leverage Multimodal Content? A Comprehensive Analysis on Multimodal Representations for Recommendation

2025-08-06 · Claudio Pomo, Matteo Attimonelli, Danilo Danese, Fedelucio Narducci 외 arxiv

Multimodal Recommender Systems aim to improve recommendation accuracy by integrating heterogeneous content, such as images and textual metadata. While effective, it remains unclear whether their gains stem from true mult…

Adversarial Multimodal Domain Transfer for Video-Level Sentiment Analysis

2022-05-11 · IEEE Access 2022 5 · Wang Yanan; Wu Jianming; Furumai Kazuaki; Wada Shinya; Kurihara Satoshi

Video-level sentiment analysis is a challenging task and requires systems to obtain discriminative multimodal representations that can capture difference in sentiments across various modalities. However, due to diverse …

Multimodal Sentiment AnalysisSentiment Analysis

EffMulti: Efficiently Modeling Complex Multimodal Interactions for Emotion Analysis

2022-12-16 · Feng Qiu, Chengyang Xie, Yu Ding, Wanzeng Kong

Humans are skilled in reading the interlocutor's emotion from multimodal signals, including spoken words, simultaneous speech, and facial expressions. It is still a challenge to effectively decode emotions from the compl…

Emotion Recognition

Evaluating Multimodal Representations on Visual Semantic Textual Similarity

2020-04-04 · Oier Lopez de Lacalle, Ander Salaberria, Aitor Soroa, Gorka Azkune 외

The combination of visual and textual representations has produced excellent results in tasks such as image captioning and visual question answering, but the inference capabilities of multimodal representations are large…

BenchmarkingImage CaptioningNatural Language InferenceQuestion Answering+3