Optimal Resource Allocation in Wireless Control Systems via Deep Policy Gradient
In wireless control systems, remote control of plants is achieved through closing of the control loop over a wireless channel. As wireless communication is noisy and subject to packet dropouts, proper allocation of limited resources, e.g. transmission power, across plants is critical for maintaining reliable operation. In this paper, we formulate the design of an optimal resource allocation policy that uses current plant states and wireless channel states to assign resources used to send control actuation information back to plants. While this problem is challenging due to its infinite dimensionality and need for explicit system model and state knowledge, we propose the use of deep reinforcement learning techniques to find neural network-based resource allocation policies. In particular, we use model-free policy gradient methods to directly learn continuous power allocation policies without knowledge of plant dynamics or communication models. Numerical simulations demonstrate the strong performance of learned policies relative to baseline resource allocation methods in settings where state information is available both with and without noise.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep Reinforcement LearningPolicy Gradient MethodsSimilar Papers 제목 키워드 기반
Model-Free Design of Control Systems over Wireless Fading Channels
Wireless control systems replace traditional wired communication with wireless networks to exchange information between actuators, plants and sensors in a control system. The noise in wireless channels renders ideal cont…
Diffusion Model Based Resource Allocation Strategy in Ultra-Reliable Wireless Networked Control Systems
Diffusion models are vastly used in generative AI, leveraging their capability to capture complex data distributions. However, their potential remains largely unexplored in the field of resource allocation in wireless ne…
Deep Reinforcement LearningDenoisingGoal-Oriented Wireless Communication Resource Allocation for Cyber-Physical Systems
The proliferation of novel industrial applications at the wireless edge, such as smart grids and vehicle networks, demands the advancement of cyber-physical systems. The performance of CPSs is closely linked to the last-…
Decision MakingDistributed OptimizationFederated LearningLearning Optimal Resource Allocations in Wireless Systems
This paper considers the design of optimal resource allocation policies in wireless communication systems which are generically modeled as a functional optimization problem with stochastic constraints. These optimization…
Large-Scale Graph Reinforcement Learning in Wireless Control Systems
Modern control systems routinely employ wireless networks to exchange information between spatially distributed plants, actuators and sensors. With wireless networks defined by random, rapidly changing transmission condi…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1