paper-with-me

홈 › Papers

LIMT: Language-Informed Multi-Task Visual World Models

2024-07-18 · Elie Aljalbout, Nikolaos Sotirakis, Patrick van der Smagt, Maximilian Karl, Nutan Chen

Most recent successes in robot reinforcement learning involve learning a specialized single-task agent. However, robots capable of performing multiple tasks can be much more valuable in real-world applications. Multi-task reinforcement learning can be very challenging due to the increased sample complexity and the potentially conflicting task objectives. Previous work on this topic is dominated by model-free approaches. The latter can be very sample inefficient even when learning specialized single-task agents. In this work, we focus on model-based multi-task reinforcement learning. We propose a method for learning multi-task visual world models, leveraging pre-trained language models to extract semantically meaningful task representations. These representations are used by the world model and policy to reason about task similarity in dynamics and behavior. Our results highlight the benefits of using language-driven task representations for world models and a clear advantage of model-based multi-task learning over the more common model-free paradigm.

📄 PDF Abstract BibTeX arXiv:2407.13466

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-Task Learningreinforcement-learningReinforcement Learning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

slimTrain -- A Stochastic Approximation Method for Training Separable Deep Neural Networks

2021-09-28 · Elizabeth Newman, Julianne Chung, Matthias Chung, Lars Ruthotto

Deep neural networks (DNNs) have shown their success as high-dimensional function approximators in many applications; however, training DNNs can be challenging in general. DNN training is commonly phrased as a stochastic…

SensitivityStochastic Optimization

LiMTR: Time Series Motion Prediction for Diverse Road Users through Multimodal Feature Integration

2024-10-21 · Camiel Oerlemans, Bram Grooten, Michiel Braat, Alaa Alassi 외

Predicting the behavior of road users accurately is crucial to enable the safe operation of autonomous vehicles in urban or densely populated areas. Therefore, there has been a growing interest in time series motion pred…

Autonomous Vehiclesmotion predictionTime Series

LiMT: A Multi-task Liver Image Benchmark Dataset

2025-11-25 · Zhe Liu, Kai Han, Siqi Ma, Yan Zhu 외 arxiv

Computer-aided diagnosis (CAD) technology can assist clinicians in evaluating liver lesions and intervening with treatment in time. Although CAD technology has advanced in recent years, the application scope of existing …

Tumor Segmentation

LimTopic: LLM-based Topic Modeling and Text Summarization for Analyzing Scientific Articles limitations

2025-03-08 · Ibrahim Al Azhar, Venkata Devesh Reddy, Hamed Alhoori, Akhil Pandey Akella

The limitations sections of scientific articles play a crucial role in highlighting the boundaries and shortcomings of research, thereby guiding future studies and improving research methods. Analyzing these limitations …

ArticlesPrompt EngineeringText Summarization

Exploring Part-Informed Visual-Language Learning for Person Re-Identification

2023-08-04 · Yin Lin, Cong Liu, Yehansen Chen, Jinshui Hu 외

Recently, visual-language learning has shown great potential in enhancing visual-based person re-identification (ReID). Existing visual-language learning-based ReID methods often focus on whole-body scale image-text feat…

Human ParsingPerson Re-Identification