paper-with-me

홈 › Papers

Position: Deployed Reinforcement Learning should be Continual

2026-06-01 · Parnian Behdin, Kevin Roice, Golnaz Mesbahi arxiv

Reinforcement Learning (RL) has received increasing attention and adoption in real-world use cases. Most of these systems follow a train-then-fix paradigm, where trained agents do not learn while interacting with the world until performance degrades and retraining becomes necessary. In this position paper, we argue that deploying an agent that is incapable of optimality, but receives an evaluative reward signal, is inherently a continual RL problem. We identify four sources of non-stationarity after deployment that necessitate never-ending learning, and highlight why the best deployed agents never stop adapting. We analyze successful examples of continual RL in the real world, and present the community with the advantages and measures to move away from the current train-then-fix paradigm.

📄 PDF Abstract BibTeX arXiv:2606.04029

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Continual Reinforcement Learning deployed in Real-life using Policy Distillation and Sim2Real Transfer

2019-06-11 · René Traoré, Hugo Caselles-Dupré, Timothée Lesort, Te Sun 외

We focus on the problem of teaching a robot to solve tasks presented sequentially, i.e., in a continual learning scenario. The robot should be able to solve all tasks it has encountered, without forgetting past tasks. We…

Continual Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Uncertainty Quantification in Continual Open-World Learning

2024-12-21 · Amanda S. Rios, Ibrahima J. Ndiour, Parual Datta, Jaroslaw Sydir 외

AI deployed in the real-world should be capable of autonomously adapting to novelties encountered after deployment. Yet, in the field of continual learning, the reliance on novelty and labeling oracles is commonplace alb…

Active LearningAI AgentContinual LearningNovelty Detection+1

CUAL: Continual Uncertainty-aware Active Learner

2024-12-12 · Amanda Rios, Ibrahima Ndiour, Parual Datta, Jerry Sydir 외

AI deployed in many real-world use cases should be capable of adapting to novelties encountered after deployment. Here, we consider a challenging, under-explored and realistic continual adaptation problem: a deployed AI …

AI Agent

Continual Learning with Adaptive Weights (CLAW)

2019-11-21 · ICLR 2020 1 · Tameem Adel, Han Zhao, Richard E. Turner

Approaches to continual learning aim to successfully learn a set of related tasks that arrive in an online manner. Recently, several frameworks have been developed which enable deep learning to be deployed in this learni…

Continual LearningTransfer LearningVariational Inference

CPPO: Continual Learning for Reinforcement Learning with Human Feedback

2024-01-16 · Conference 2024 1 · Han Zhang, Yu Lei, Lin Gui, Min Yang 외

The approach of Reinforcement Learning from Human Feedback (RLHF) is widely used for enhancing pre-trained Language Models (LM), enabling them to better align with human preferences. Existing RLHF-based LMs however req…

Continual Learningreinforcement-learningReinforcement Learning