paper-with-me

Papers

Distributed Multitask Reinforcement Learning with Quadratic Convergence

2018-12-01 · NeurIPS 2018 12 · Rasul Tutunov, Dongho Kim, Haitham Bou Ammar

Multitask reinforcement learning (MTRL) suffers from scalability issues when the number of tasks or trajectories grows large. The main reason behind this drawback is the reliance on centeralised solutions. Recent methods exploited the connection between MTRL and general consensus to propose scalable solutions. These methods, however, suffer from two drawbacks. First, they rely on predefined objectives, and, second, exhibit linear convergence guarantees. In this paper, we improve over state-of-the-art by deriving multitask reinforcement learning from a variational inference perspective. We then propose a novel distributed solver for MTRL with quadratic convergence guarantees.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Variational Inference

Similar Papers 제목 키워드 기반

Fully Distributed Actor-Critic Architecture for Multitask Deep Reinforcement Learning

2021-10-23 · Sergio Valcarcel Macua, Ian Davies, Aleksi Tukiainen, Enrique Munoz de Cote

We propose a fully distributed actor-critic architecture, named Diff-DAC, with application to multitask reinforcement learning (MRL). During the learning process, agents communicate their value and policy parameters to t…

continuous-controlContinuous ControlDeep Reinforcement Learningreinforcement-learning+2

On Centralized and Distributed Mirror Descent: Convergence Analysis Using Quadratic Constraints

2021-05-29 · Youbang Sun, Mahyar Fazlyab, Shahin Shahrampour

Mirror descent (MD) is a powerful first-order optimization technique that subsumes several optimization algorithms including gradient descent (GD). In this work, we develop a semi-definite programming (SDP) framework to …

Policy Gradient Methods for Discrete Time Linear Quadratic Regulator With Random Parameters

2023-03-29 · Deyue Li

This paper studies an infinite horizon optimal control problem for discrete-time linear system and quadratic criteria, both with random parameters which are independent and identically distributed with respect to time. I…

Policy Gradient Methodsreinforcement-learning

Robust Multitask Diffusion Normalized M-estimate Subband Adaptive Filtering Algorithm Over Adaptive Networks

2022-10-20 · Wenjing Xu, Haiquan Zhao, Shaohui Lv

In recent years, the multitask diffusion least mean square (MD-LMS) algorithm has been extensively applied in the distributed parameter estimation and target tracking of multitask network. However, its performance is mai…

parameter estimation

Diff-DAC: Distributed Actor-Critic for Average Multitask Deep Reinforcement Learning

2017-10-28 · Sergio Valcarcel Macua, Aleksi Tukiainen, Daniel García-Ocaña Hernández, David Baldazo 외

We propose a fully distributed actor-critic algorithm approximated by deep neural networks, named \textit{Diff-DAC}, with application to single-task and to average multitask reinforcement learning (MRL). Each agent has a…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)