paper-with-me

홈 › Papers

The association problem in wireless networks: a Policy Gradient Reinforcement Learning approach

2013-06-11 · Richard Combes, Ilham El Bouloumi, Stephane Senecal, Zwi Altman

The purpose of this paper is to develop a self-optimized association algorithm based on PGRL (Policy Gradient Reinforcement Learning), which is both scalable, stable and robust. The term robust means that performance degradation in the learning phase should be forbidden or limited to predefined thresholds. The algorithm is model-free (as opposed to Value Iteration) and robust (as opposed to Q-Learning). The association problem is modeled as a Markov Decision Process (MDP). The policy space is parameterized. The parameterized family of policies is then used as expert knowledge for the PGRL. The PGRL converges towards a local optimum and the average cost decreases monotonically during the learning process. The properties of the solution make it a good candidate for practical implementation. Furthermore, the robustness property allows to use the PGRL algorithm in an "always-on" learning mode.

📄 PDF Abstract BibTeX arXiv:1306.2554

Code (0)

등록된 구현이 없습니다.

Tasks

Q-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Meta-Reinforcement Learning for Reliable Communication in THz/VLC Wireless VR Networks

2021-01-29 · Yining Wang, Mingzhe Chen, Zhaohui Yang, Walid Saad 외

In this paper, the problem of enhancing the quality of virtual reality (VR) services is studied for an indoor terahertz (THz)/visible light communication (VLC) wireless network. In the studied model, small base stations …

Meta-LearningMeta Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Privacy-Preserving Joint Edge Association and Power Optimization for the Internet of Vehicles via Federated Multi-Agent Reinforcement Learning

2023-01-26 · Yan Lin, Jinming Bao, Yijin Zhang, Jun Li 외

Proactive edge association is capable of improving wireless connectivity at the cost of increased handover (HO) frequency and energy consumption, while relying on a large amount of private information sharing required fo…

Decision MakingMulti-agent Reinforcement LearningPrivacy Preserving

Scalable and Sample Efficient Distributed Policy Gradient Algorithms in Multi-Agent Networked Systems

2022-12-13 · Xin Liu, Honghao Wei, Lei Ying

This paper studies a class of multi-agent reinforcement learning (MARL) problems where the reward that an agent receives depends on the states of other agents, but the next state only depends on the agent's own current s…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Optimal Resource Allocation in Wireless Control Systems via Deep Policy Gradient

2019-10-25

In wireless control systems, remote control of plants is achieved through closing of the control loop over a wireless channel. As wireless communication is noisy and subject to packet dropouts, proper allocation of limit…

Deep Reinforcement LearningPolicy Gradient Methods

Deep Reinforcement Learning for Wireless Scheduling in Distributed Networked Control

2021-09-26 · Gaoyang Pang, Kang Huang, Daniel E. Quevedo, Branka Vucetic 외

We consider a joint uplink and downlink scheduling problem of a fully distributed wireless networked control system (WNCS) with a limited number of frequency channels. Using elements of stochastic systems theory, we deri…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Scheduling