paper-with-me

Papers

Non-Stationary Policy Learning for Multi-Timescale Multi-Agent Reinforcement Learning

2023-07-17 · Patrick Emami, Xiangyu Zhang, David Biagioni, Ahmed S. Zamzam

In multi-timescale multi-agent reinforcement learning (MARL), agents interact across different timescales. In general, policies for time-dependent behaviors, such as those induced by multiple timescales, are non-stationary. Learning non-stationary policies is challenging and typically requires sophisticated or inefficient algorithms. Motivated by the prevalence of this control problem in real-world complex systems, we introduce a simple framework for learning non-stationary policies for multi-timescale MARL. Our approach uses available information about agent timescales to define a periodic time encoding. In detail, we theoretically demonstrate that the effects of non-stationarity introduced by multiple timescales can be learned by a periodic multi-agent policy. To learn such policies, we propose a policy gradient algorithm that parameterizes the actor and critic with phase-functioned neural networks, which provide an inductive bias for periodicity. The framework's ability to effectively learn multi-timescale policies is validated on a gridworld and building energy management environment.

📄 PDF Abstract BibTeX arXiv:2307.08794

Code (0)

등록된 구현이 없습니다.

Tasks

energy managementInductive BiasManagementMulti-agent Reinforcement Learningreinforcement-learning

Similar Papers 제목 키워드 기반

Bi-level Off-policy Reinforcement Learning for Volt/VAR Control Involving Continuous and Discrete Devices

2021-04-13 · Haotian Liu, Wenchuan Wu

In Volt/Var control (VVC) of active distribution networks(ADNs), both slow timescale discrete devices (STDDs) and fast timescale continuous devices (FTCDs) are involved. The STDDs such as on-load tap changers (OLTC) and …

Reinforcement Learning (RL)

Self-Monitoring Benefits from Structural Integration: Lessons from Metacognition in Continuous-Time Multi-Timescale Agents

2026-04-13 · Ying Xie arxiv

Self-monitoring capabilities -- metacognition, self-prediction, and subjective duration -- are often proposed as useful additions to reinforcement learning agents. But do they actually help? We investigate this question …

Reinforcement Learning

Dealing With Non-stationarity in Decentralized Cooperative Multi-Agent Deep Reinforcement Learning via Multi-Timescale Learning

2023-02-06 · Hadi Nekoei, Akilesh Badrinaaraayanan, Amit Sinha, Mohammad Amini 외

Decentralized cooperative multi-agent deep reinforcement learning (MARL) can be a versatile learning framework, particularly in scenarios where centralized training is either not possible or not practical. One of the cri…

Deep Reinforcement Learning

Continual Reinforcement Learning with Multi-Timescale Replay

2020-04-16 · Christos Kaplanis, Claudia Clopath, Murray Shanahan

In this paper, we propose a multi-timescale replay (MTR) buffer for improving continual learning in RL agents faced with environments that are changing continuously over time at timescales that are unknown to the agent. …

Continual Learningcontinuous-controlContinuous Controlreinforcement-learning+2

A Policy Gradient Algorithm for Learning to Learn in Multiagent Reinforcement Learning

2020-10-31 · Dong-Ki Kim, Miao Liu, Matthew Riemer, Chuangchuang Sun 외

A fundamental challenge in multiagent reinforcement learning is to learn beneficial behaviors in a shared environment with other simultaneously learning agents. In particular, each agent perceives the environment as effe…

reinforcement-learningReinforcement Learning (RL)