Off-policy Learning for Remote Electrical Tilt Optimization
We address the problem of Remote Electrical Tilt (RET) optimization using off-policy Contextual Multi-Armed-Bandit (CMAB) techniques. The goal in RET optimization is to control the orientation of the vertical tilt angle of the antenna to optimize Key Performance Indicators (KPIs) representing the Quality of Service (QoS) perceived by the users in cellular networks. Learning an improved tilt update policy is hard. On the one hand, coming up with a new policy in an online manner in a real network requires exploring tilt updates that have never been used before, and is operationally too risky. On the other hand, devising this policy via simulations suffers from the simulation-to-reality gap. In this paper, we circumvent these issues by learning an improved policy in an offline manner using existing data collected on real networks. We formulate the problem of devising such a policy using the off-policy CMAB framework. We propose CMAB learning algorithms to extract optimal tilt update policies from the data. We train and evaluate these policies on real-world 4G Long Term Evolution (LTE) cellular network data. Our policies show consistent improvements over the rule-based logging policy used to collect the data.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Remote Electrical Tilt Optimization via Safe Reinforcement Learning
Remote Electrical Tilt (RET) optimization is an efficient method for adjusting the vertical tilt angle of Base Stations (BSs) antennas in order to optimize Key Performance Indicators (KPIs) of the network. Reinforcement …
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement Learning+1Learning Optimal Antenna Tilt Control Policies: A Contextual Linear Bandit Approach
Controlling antenna tilts in cellular networks is imperative to reach an efficient trade-off between network coverage and capacity. In this paper, we devise algorithms learning optimal tilt control policies from existing…
Active LearningA Safe Reinforcement Learning Architecture for Antenna Tilt Optimisation
Safe interaction with the environment is one of the most challenging aspects of Reinforcement Learning (RL) when applied to real-world problems. This is particularly important when unsafe actions have a high or irreversi…
Managementreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1Safe RAN control: A Symbolic Reinforcement Learning Approach
In this paper, we present a Symbolic Reinforcement Learning (SRL) based architecture for safety control of Radio Access Network (RAN) applications. In particular, we provide a purely automated procedure in which a user c…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Self-LearningMulti-Agent Reinforcement Learning with Common Policy for Antenna Tilt Optimization
This paper presents a method for optimizing wireless networks by adjusting cell parameters that affect both the performance of the cell being optimized and the surrounding cells. The method uses multiple reinforcement le…
Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)