paper-with-me

홈 › Papers

Model Based Residual Policy Learning with Applications to Antenna Control

2022-11-16 · Viktor Eriksson Möllerstedt, Alessio Russo, Maxime Bouton

Non-differentiable controllers and rule-based policies are widely used for controlling real systems such as telecommunication networks and robots. Specifically, parameters of mobile network base station antennas can be dynamically configured by these policies to improve users coverage and quality of service. Motivated by the antenna tilt control problem, we introduce Model-Based Residual Policy Learning (MBRPL), a practical reinforcement learning (RL) method. MBRPL enhances existing policies through a model-based approach, leading to improved sample efficiency and a decreased number of interactions with the actual environment when compared to off-the-shelf RL methods.To the best of our knowledge, this is the first paper that examines a model-based approach for antenna control. Experimental results reveal that our method delivers strong initial performance while improving sample efficiency over previous RL methods, which is one step towards deploying these algorithms in real networks.

📄 PDF Abstract BibTeX arXiv:2211.08796

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Golden Queue Managers 설명 없음
BASE 설명 없음

Similar Papers 제목 키워드 기반

Learning Optimal Antenna Tilt Control Policies: A Contextual Linear Bandit Approach

2022-01-06 · Filippo Vannella, Alexandre Proutiere, Yassir Jedra, Jaeseong Jeong

Controlling antenna tilts in cellular networks is imperative to reach an efficient trade-off between network coverage and capacity. In this paper, we devise algorithms learning optimal tilt control policies from existing…

Active Learning

Residual Deep Reinforcement Learning for Inverter-based Volt-Var Control

2024-08-13 · Qiong Liu, Ye Guo, Lirong Deng, Haotian Liu 외

A residual deep reinforcement learning (RDRL) approach is proposed by integrating DRL with model-based optimization for inverter-based volt-var control in active distribution networks when the accurate power flow model i…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning

Efficient Real-World Autonomous Racing via Attenuated Residual Policy Optimization

2026-03-13 · Raphael Trumpp, Denis Hoornaert, Mirco Theile, Marco Caccamo arxiv

Residual policy learning (RPL), in which a learned policy refines a static base policy using deep reinforcement learning (DRL), has shown strong performance across various robotic applications. Its effectiveness is parti…

Reinforcement Learning

Bellman Residual Minimization for Control: Geometry, Stationarity, and Convergence

2026-01-26 · Donghwan Lee, Hyukjun Yang arxiv

Markov decision problems are most commonly solved via dynamic programming. Another approach is Bellman residual minimization, which directly minimizes the squared Bellman residual objective function. However, compared to…

Reinforcement Learning

Multi-residual Mixture of Experts Learning for Cooperative Control in Multi-vehicle Systems

2025-07-14 · Vindula Jayawardana, Sirui Li, Yashar Farid, Cathy Wu arxiv

Autonomous vehicles (AVs) are becoming increasingly popular, with their applications now extending beyond just a mode of transportation to serving as mobile actuators of a traffic flow to control flow dynamics. This cont…

Reinforcement LearningAutonomous Vehicles