paper-with-me

Papers

Comparative Field Deployment of Reinforcement Learning and Model Predictive Control for Residential HVAC

2025-10-01 · Ozan Baris Mulayim, Elias N. Pergantis, Levi D. Reyes Premer, Bingqing Chen, Guannan Qu, Kevin J. Kircher, Mario Bergés arxiv

Model Predictive Control (MPC) has demonstrated significant performance improvements over today's control methods for residential Heating, Ventilation, and Air Conditioning (HVAC), but deploying MPC often requires substantial engineering effort. Reinforcement Learning (RL) may offer comparable performance with easier deployment, but its practical application for residential HVAC remains largely undemonstrated, leaving open questions related to occupant comfort and data requirements. To investigate these issues, we deployed one MPC variant and one model-based RL variant for one month each in an occupied house in a cold climate. The controllers adjusted an air-to-air heat pump's thermostat temperature setpoint based on measurements of the indoor temperature and the electric power used for heating. Relative to constant-setpoint operation, MPC saved 18.1\% (95\% confidence interval: 4.4 to 30.9\%) of weather-normalized heat pump energy and RL saved 20.9\% (2.6 to 38.3\%). MPC maintained acceptable occupant comfort. RL kept the house cooler, particularly during an initial adaptation phase, leading to three reports of occupant discomfort. The two algorithms had similar data requirements. We estimate that for a fresh deployment in another house, RL would take about one-third less engineering effort than MPC. While RL reduces deployment effort, it faces difficulties related to safe controller initialization and to mismatches between the modeled and true state and action spaces.

📄 PDF Abstract BibTeX arXiv:2510.01475

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Reinforcement Learning with Ensemble Model Predictive Safety Certification

2024-02-06 · Sven Gronauer, Tom Haider, Felippe Schmoeller da Roza, Klaus Diepold

Reinforcement learning algorithms need exploration to learn. However, unsupervised exploration prevents the deployment of such algorithms on safety-critical tasks and limits real-world deployment. In this paper, we propo…

Deep Reinforcement LearningmodelModel Predictive Controlreinforcement-learning+1

Lessons learned from field demonstrations of model predictive control and reinforcement learning for residential and commercial HVAC: A review

2025-03-06 · Arash J. Khabbazi, Elias N. Pergantis, Levi D. Reyes Premer, Panagiotis Papageorgiou 외

A large body of simulation research suggests that model predictive control (MPC) and reinforcement learning (RL) for heating, ventilation, and air-conditioning (HVAC) in residential and commercial buildings could reduce …

Model Predictive ControlReinforcement Learning (RL)

Model-free Vehicle Rollover Prevention: A Data-driven Predictive Control Approach

2025-03-26 · Mohammad R. Hajidavalloo, Kaixiang Zhang, Vaibhav Srivastava, Zhaojian Li

Vehicle rollovers pose a significant safety risk and account for a disproportionately high number of fatalities in road accidents. This paper addresses the challenge of rollover prevention using Data-EnablEd Predictive C…

Computational EfficiencyDimensionality ReductionModel Predictive Control

Deep Reinforcement Learning for Optimizing Inverter Control: Fixed and Adaptive Gain Tuning Strategies for Power System Stability

2024-11-03 · Shuvangkar Chandra Das, Tuyen Vu, Deepak Ramasubramanian, Evangelos Farantatos 외

This paper presents novel methods for tuning inverter controller gains using deep reinforcement learning (DRL). A Simulink-developed inverter model is converted into a dynamic link library (DLL) and integrated with a Pyt…

Deep Reinforcement Learning

Safe Reinforcement Learning using Data-Driven Predictive Control

2022-11-20 · Mahmoud Selim, Amr Alanwar, M. Watheq El-Kharashi, Hazem M. Abbas 외

Reinforcement learning (RL) algorithms can achieve state-of-the-art performance in decision-making and continuous control tasks. However, applying RL algorithms on safety-critical systems still needs to be well justified…

continuous-controlContinuous ControlDecision Makingreinforcement-learning+3