A Safe Reinforcement Learning Architecture for Antenna Tilt Optimisation
Safe interaction with the environment is one of the most challenging aspects of Reinforcement Learning (RL) when applied to real-world problems. This is particularly important when unsafe actions have a high or irreversible negative impact on the environment. In the context of network management operations, Remote Electrical Tilt (RET) optimisation is a safety-critical application in which exploratory modifications of antenna tilt angles of base stations can cause significant performance degradation in the network. In this paper, we propose a modular Safe Reinforcement Learning (SRL) architecture which is then used to address the RET optimisation in cellular networks. In this approach, a safety shield continuously benchmarks the performance of RL agents against safe baselines, and determines safe antenna tilt updates to be performed on the network. Our results demonstrate improved performance of the SRL agent over the baseline while ensuring the safety of the performed actions.
Code (0)
등록된 구현이 없습니다.
Tasks
Managementreinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement LearningSimilar Papers 제목 키워드 기반
Remote Electrical Tilt Optimization via Safe Reinforcement Learning
Remote Electrical Tilt (RET) optimization is an efficient method for adjusting the vertical tilt angle of Base Stations (BSs) antennas in order to optimize Key Performance Indicators (KPIs) of the network. Reinforcement …
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement Learning+1Safe RAN control: A Symbolic Reinforcement Learning Approach
In this paper, we present a Symbolic Reinforcement Learning (SRL) based architecture for safety control of Radio Access Network (RAN) applications. In particular, we provide a purely automated procedure in which a user c…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Self-LearningBase Station Antenna Uptilt Optimization for Cellular-Connected Drone Corridors
The concept of drone corridors is recently getting more attention to enable connected, safe, and secure flight zones in the national airspace. To support beyond visual line of sight (BVLOS) operations of aerial vehicles …
Multi-agent Reinforcement Learning with Graph Q-Networks for Antenna Tuning
Future generations of mobile networks are expected to contain more and more antennas with growing complexity and more parameters. Optimizing these parameters is necessary for ensuring the good performance of the network.…
Graph Neural NetworkMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1Ensuring Reliable Connectivity to Cellular-Connected UAVs with Up-tilted Antennas and Interference Coordination
To integrate unmanned aerial vehicles (UAVs) in future large-scale deployments, a new wireless communication paradigm, namely, the cellular-connected UAV has recently attracted interest. However, the line-of-sight domina…