paper-with-me

홈 › Papers

Offline-Online Reinforcement Learning for Energy Pricing in Office Demand Response: Lowering Energy and Data Costs

2021-08-14 · Doseok Jang, Lucas Spangher, Manan Khattar, Utkarsha Agwan, Selvaprabuh Nadarajah, Costas Spanos

Our team is proposing to run a full-scale energy demand response experiment in an office building. Although this is an exciting endeavor which will provide value to the community, collecting training data for the reinforcement learning agent is costly and will be limited. In this work, we examine how offline training can be leveraged to minimize data costs (accelerate convergence) and program implementation costs. We present two approaches to doing so: pretraining our model to warm start the experiment with simulated tasks, and using a planning model trained to simulate the real world's rewards to the agent. We present results that demonstrate the utility of offline reinforcement learning to efficient price-setting in the energy demand response problem.

📄 PDF Abstract BibTeX arXiv:2108.06594

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Offline Deep Reinforcement Learning for Dynamic Pricing of Consumer Credit

2022-03-06 · Raad Khraishi, Ramin Okhrati

We introduce a method for pricing consumer credit using recent advances in offline deep reinforcement learning. This approach relies on a static dataset and requires no assumptions on the functional form of demand. Using…

Deep Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning+1

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

2026-05-12 · Jose E. Aguilar Escamilla, Lingdong Zhou, Xiangqi Zhu, Huazheng Wang arxiv

Extreme weather and volatile wholesale electricity markets expose residential consumers to catastrophic financial risks, yet demand response at the distribution level remains an underutilized tool for grid flexibility an…

Reinforcement Learning

Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning

2024-07-17 · Xu-Hui Liu, Tian-Shuo Liu, Shengyi Jiang, Ruifeng Chen 외

Combining offline and online reinforcement learning (RL) techniques is indeed crucial for achieving efficient and safe learning where data acquisition is expensive. Existing methods replay offline data directly in the on…

MuJoCoreinforcement-learningReinforcement LearningReinforcement Learning (RL)

How does online shopping affect offline price sensitivity?

2025-06-18 · Shirsho Biswas, Hema Yoganarasimhan, Haonan Zhang

The rapid rise of e-commerce has transformed consumer behavior, prompting questions about how online adoption influences offline shopping. We examine whether consumers who adopt online shopping with a retailer become mor…

counterfactualSensitivity

AIGP: An LLM-Based Framework for Long-Term Value Alignment in E-Commerce Pricing

2026-06-25 · Chennan Ma, Yanning Zhang, Siqi Hong, Xiuchong Wang 외 arxiv

Traditional dynamic pricing models in large-scale e-commerce suffer from limited interpretability, poor utilization of unstructured information, and misalignment with long-term business objectives such as cumulative Gros…

Knowledge DistillationReinforcement Learning