paper-with-me

홈 › Papers

Generalized Policy Learning for Smart Grids: FL TRPO Approach

2024-03-27 · Yunxiang Li, Nicolas Mauricio Cuadrado, Samuel Horváth, Martin Takáč

The smart grid domain requires bolstering the capabilities of existing energy management systems; Federated Learning (FL) aligns with this goal as it demonstrates a remarkable ability to train models on heterogeneous datasets while maintaining data privacy, making it suitable for smart grid applications, which often involve disparate data distributions and interdependencies among features that hinder the suitability of linear models. This paper introduces a framework that combines FL with a Trust Region Policy Optimization (FL TRPO) aiming to reduce energy-associated emissions and costs. Our approach reveals latent interconnections and employs personalized encoding methods to capture unique insights, understanding the relationships between features and optimal strategies, allowing our model to generalize to previously unseen data. Experimental results validate the robustness of our approach, affirming its proficiency in effectively learning policy models for smart grid challenges.

📄 PDF Abstract BibTeX arXiv:2403.18439

Code (0)

등록된 구현이 없습니다.

Tasks

energy managementFederated LearningManagement

Similar Papers 제목 키워드 기반

Generalizing in Net-Zero Microgrids: A Study with Federated PPO and TRPO

2024-12-30 · Nicolas M Cuadrado Avila, Samuel Horváth, Martin Takáč

This work addresses the challenge of optimal energy management in microgrids through a collaborative and privacy-preserving framework. We propose the FedTRPO methodology, which integrates Federated Learning (FL) and Trus…

energy managementFederated LearningManagementPrivacy Preserving

Trust-Region-Free Policy Optimization for Stochastic Policies

2023-02-15 · Mingfei Sun, Benjamin Ellis, Anuj Mahajan, Sam Devlin 외

Trust Region Policy Optimization (TRPO) is an iterative method that simultaneously maximizes a surrogate objective and enforces a trust region constraint over consecutive policies in each iteration. The combination of th…

EnTRPO: Trust Region Policy Optimization Method with Entropy Regularization

2021-10-26 · Sahar Roostaie, Mohammad Mehdi Ebadzadeh

Trust Region Policy Optimization (TRPO) is a popular and empirically successful policy search algorithm in reinforcement learning (RL). It iteratively solved the surrogate problem which restricts consecutive policies to …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Cooperative Dispatch of Microgrids Community Using Risk-Sensitive Reinforcement Learning with Monotonously Improved Performance

2023-10-17 · Ziqing Zhu, Xiang Gao, Siqi Bu, Ka Wing Chan 외

The integration of individual microgrids (MGs) into Microgrid Clusters (MGCs) significantly improves the reliability and flexibility of energy supply, through resource sharing and ensuring backup during outages. The disp…

Adaptive Trust Region Policy Optimization: Global Convergence and Faster Rates for Regularized MDPs

2019-09-06 · Lior Shani, Yonathan Efroni, Shie Mannor

Trust region policy optimization (TRPO) is a popular and empirically successful policy search algorithm in Reinforcement Learning (RL) in which a surrogate problem, that restricts consecutive policies to be 'close' to on…

Reinforcement LearningReinforcement Learning (RL)