paper-with-me

홈 › Papers

Combined Peak Reduction and Self-Consumption Using Proximal Policy Optimization

2022-11-27 · Thijs Peirelinck, Chris Hermans, Fred Spiessens, Geert Deconinck

Residential demand response programs aim to activate demand flexibility at the household level. In recent years, reinforcement learning (RL) has gained significant attention for these type of applications. A major challenge of RL algorithms is data efficiency. New RL algorithms, such as proximal policy optimisation (PPO), have tried to increase data efficiency. Additionally, combining RL with transfer learning has been proposed in an effort to mitigate this challenge. In this work, we further improve upon state-of-the-art transfer learning performance by incorporating demand response domain knowledge into the learning pipeline. We evaluate our approach on a demand response use case where peak shaving and self-consumption is incentivised by means of a capacity tariff. We show our adapted version of PPO, combined with transfer learning, reduces cost by 14.51% compared to a regular hysteresis controller and by 6.68% compared to traditional PPO.

📄 PDF Abstract BibTeX arXiv:2211.14831

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)Transfer Learning

Methods 이 논문이 사용한 방법론

Entropy Regularization 설명 없음
PPO Proximal Policy Optimization, or PPO, is a policy gradient method for reinforcement learning. The motivation was to have an algorithm with the data efficiency and reliable…

Similar Papers 제목 키워드 기반

PReS: Power Peak Reduction by Real-time Scheduling for Urban Railway Transit

2019-02-21

Railway transportation is one of the most popular options for Urban Massive Transportation Systems (UMTS) because of many attractive features. A robust electric power supply is essential to enable normal operation. Howev…

Scheduling

Self-supervised Speaker Recognition Training Using Human-Machine Dialogues

2022-02-07 · Metehan Cekic, Ruirui Li, Zeya Chen, Yuguang Yang 외

Speaker recognition, recognizing speaker identities based on voice alone, enables important downstream applications, such as personalization and authentication. Learning speaker representations, in the context of supervi…

Contrastive LearningSpeaker Recognition

Self-attention encoding and pooling for speaker recognition

2020-08-03 · Pooyan Safari, Miquel India, Javier Hernando

The computing power of mobile devices limits the end-user applications in terms of storage size, processing, memory and energy consumption. These limitations motivate researchers for the design of more efficient deep mod…

Speaker RecognitionSpeaker VerificationText-Independent Speaker Verification

Risk-aware Scheduling and Dispatch of Flexibility Events in Buildings

2023-11-09 · Paul Scharnhorst, Baptiste Schubnel, Rafael E. Carrillo, Pierre-Jean Alet 외

Residential and commercial buildings, equipped with systems such as heat pumps (HPs), hot water tanks, or stationary energy storage, have a large potential to offer their consumption flexibility as grid services. In this…

Scheduling

Learning to Unscramble Feynman Loop Integrals with SAILIR

2026-04-06 · David Shih arxiv

Integration-by-parts (IBP) reduction of Feynman integrals to master integrals is a key computational bottleneck in precision calculations in high-energy physics. Traditional approaches based on the Laporta algorithm requ…