paper-with-me

Papers Meta Reinforcement Learning

“Meta Reinforcement Learning” 태그가 달린 논문 278편 · 필터 해제

Meta-Reinforcement Learning for Fast and Data-Efficient Spectrum Allocation in Dynamic Wireless Networks

2025-07-13 · Oluwaseyi Giwa, Tobi Awodunmila, Muhammad Ahmed Mohsin, Ahsan Bilal 외

The dynamic allocation of spectrum in 5G / 6G networks is critical to efficient resource utilization. However, applying traditional deep reinforcement learning (DRL) is often infeasible due to its immense sample complexi…

Deep Reinforcement LearningFairnessMeta-LearningMeta Reinforcement Learning

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning

2025-06-24 · Menglong Zhang, Fuyuan Qian

Meta-reinforcement learning requires utilizing prior task distribution information obtained during exploration to rapidly adapt to unknown tasks. The efficiency of an agent's exploration hinges on accurately identifying …

Meta Reinforcement LearningMuJoCo

Scaling Algorithm Distillation for Continuous Control with Mamba

2025-06-16 · Samuel Beaussant, Mehdi Mounsif

Algorithm Distillation (AD) was recently proposed as a new approach to perform In-Context Reinforcement Learning (ICRL) by modeling across-episodic training histories autoregressively with a causal transformer model. How…

continuous-controlContinuous ControlIn-Context Reinforcement LearningMamba+3

Unsupervised Meta-Testing with Conditional Neural Processes for Hybrid Meta-Reinforcement Learning

2025-06-04 · Suzan Ece Ada, Emre Ugur

We introduce Unsupervised Meta-Testing with Conditional Neural Processes (UMCNP), a novel hybrid few-shot meta-reinforcement learning (meta-RL) method that uniquely combines, yet distinctly separates, parameterized polic…

continuous-controlContinuous ControlMeta Reinforcement Learning

Bayesian Meta-Reinforcement Learning with Laplace Variational Recurrent Networks

2025-05-24 · Joery A. de Vries, Jinke He, Mathijs M. de Weerdt, Matthijs T. J. Spaan

Meta-reinforcement learning trains a single reinforcement learning agent on a distribution of tasks to quickly generalize to new tasks outside of the training set at test time. From a Bayesian perspective, one can interp…

Meta Reinforcement Learningreinforcement-learningReinforcement LearningVariational Inference

Meta-reinforcement learning with minimum attention

2025-05-22 · Pilhwa Lee, Shashank Gupta

Minimum attention applies the least action principle in the changes of control concerning state and time, first proposed by Brockett. The involved regularization is highly relevant in emulating biological control, such a…

Meta-LearningMeta Reinforcement Learningreinforcement-learningReinforcement Learning+1

Meta-World+: An Improved, Standardized, RL Benchmark

2025-05-16 · Reginald McLean, Evangelos Chatzaroulas, Luc McCutcheon, Frank Röder 외

Meta-World is widely used for evaluating multi-task and meta-reinforcement learning agents, which are challenged to master diverse skills simultaneously. Since its introduction however, there have been numerous undocumen…

Meta Reinforcement Learningreinforcement-learningReinforcement Learning

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments

2025-04-27 · Yun Qu, Qi Cheems Wang, Yixiu Mao, Yiqin Lv 외

Task robust adaptation is a long-standing pursuit in sequential decision-making. Some risk-averse strategies, e.g., the conditional value-at-risk principle, are incorporated in domain randomization or meta reinforcement …

Decision MakingDiversityMeta Reinforcement LearningSequential Decision Making

InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning

2025-04-17 · Zheng Wang, Shu Xian Teo, Jun Jie Chew, Wei Shi

Recent advancements in large language models (LLMs) have enabled their use as agents for planning complex tasks. Existing methods typically rely on a thought-action-observation (TAO) process to enhance LLM performance, b…

Meta-LearningMeta Reinforcement LearningRAGreinforcement-learning+4

Embodied World Models Emerge from Navigational Task in Open-Ended Environments

2025-04-15 · Li Jin, Liu Jia

Spatial reasoning in partially observable environments has often been approached through passive predictive models, yet theories of embodied cognition suggest that genuinely useful representations arise only when percept…

Meta Reinforcement LearningSpatial Reasoning

UAS Visual Navigation in Large and Unseen Environments via a Meta Agent

2025-03-20 · Yuci Han, Charles Toth, Alper Yilmaz

The aim of this work is to develop an approach that enables Unmanned Aerial System (UAS) to efficiently learn to navigate in large-scale urban environments and transfer their acquired expertise to novel environments. To …

Incremental LearningMeta Reinforcement LearningNavigatePhilosophy+4

Meta-Reinforcement Learning with Discrete World Models for Adaptive Load Balancing

2025-03-11 · Cameron Redovian

We integrate a meta-reinforcement learning algorithm with the DreamerV3 architecture to improve load balancing in operating systems. This approach enables rapid adaptation to dynamic workloads with minimal retraining, ou…

ManagementMeta Reinforcement Learningreinforcement-learningReinforcement Learning

Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning

2025-03-10 · Yuxiao Qu, Matthew Y. R. Yang, Amrith Setlur, Lewis Tunstall 외

Training models to effectively use test-time compute is crucial for improving the reasoning performance of LLMs. Current methods mostly do so via fine-tuning on search traces or running RL with 0/1 outcome reward, but do…

MathMeta Reinforcement LearningReinforcement Learning (RL)

Teleology-Driven Affective Computing: A Causal Framework for Sustained Well-Being

2025-02-24 · Bin Yin, Chong-Yi Liu, Liya Fu, Jinkun Zhang

Affective computing has made significant strides in emotion recognition and generation, yet current approaches mainly focus on short-term pattern recognition and lack a comprehensive framework to guide affective agents t…

Emotion RecognitionMeta Reinforcement Learning

PRISM: A Robust Framework for Skill-based Meta-Reinforcement Learning with Noisy Demonstrations

2025-02-06 · Sanghyeon Lee, Sangjun Bae, Yisak Park, Seungyul Han

Meta-reinforcement learning (Meta-RL) facilitates rapid adaptation to unseen tasks but faces challenges in long-horizon environments. Skill-based approaches tackle this by decomposing state-action sequences into reusable…

Decision MakingMeta Reinforcement Learning

Task-Aware Virtual Training: Enhancing Generalization in Meta-Reinforcement Learning for Out-of-Distribution Tasks

2025-02-05 · Jeongmo Kim, Yisak Park, Minung Kim, Seungyul Han

Meta reinforcement learning aims to develop policies that generalize to unseen tasks sampled from a task distribution. While context-based meta-RL methods improve task representation using task latents, they often strugg…

Meta Reinforcement LearningMuJoCoRepresentation Learning

Coreset-Based Task Selection for Sample-Efficient Meta-Reinforcement Learning

2025-02-04 · Donglin Zhan, Leonardo F. Toso, James Anderson

We study task selection to enhance sample efficiency in model-agnostic meta-reinforcement learning (MAML-RL). Traditional meta-RL typically assumes that all available tasks are equally important, which can lead to task r…

Meta Reinforcement Learning

Toward Task Generalization via Memory Augmentation in Meta-Reinforcement Learning

2025-02-03 · Kaixi Bao, Chenhao Li, Yarden As, Andreas Krause 외

Agents trained via reinforcement learning (RL) often struggle to perform well on tasks that differ from those encountered during training. This limitation presents a challenge to the broader deployment of RL in diverse a…

Meta Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

TIMRL: A Novel Meta-Reinforcement Learning Framework for Non-Stationary and Multi-Task Environments

2025-01-13 · Chenyang Qi, Huiping Li, Panfeng Huang

In recent years, meta-reinforcement learning (meta-RL) algorithm has been proposed to improve sample efficiency in the field of decision-making and control, enabling agents to learn new knowledge from a small number of s…

Decision MakingMeta Reinforcement LearningMuJoCoreinforcement-learning+1

Hierarchical Multi-agent Meta-Reinforcement Learning for Cross-channel Bidding

2024-12-26 · Shenghong He, Chao Yu

Real-time bidding (RTB) plays a pivotal role in online advertising ecosystems. Advertisers employ strategic bidding to optimize their advertising impact while adhering to various financial constraints, such as the return…

global-optimizationMeta Reinforcement LearningMulti-agent Reinforcement Learningreinforcement-learning+1
1–20 / 278 다음 →