paper-with-me

홈 › Papers

BEAVER: Building Environments with Assessable Variation for Evaluating Multi-Objective Reinforcement Learning

2025-07-10 · Ruohong Liu, Jack Umenberger, Yize Chen arxiv

Recent years have seen significant advancements in designing reinforcement learning (RL)-based agents for building energy management. While individual success is observed in simulated or controlled environments, the scalability of RL approaches in terms of efficiency and generalization across building dynamics and operational scenarios remains an open question. In this work, we formally characterize the generalization space for the cross-environment, multi-objective building energy management task, and formulate the multi-objective contextual RL problem. Such a formulation helps understand the challenges of transferring learned policies across varied operational contexts such as climate and heat convection dynamics under multiple control objectives such as comfort level and energy consumption. We provide a principled way to parameterize such contextual information in realistic building RL environments, and construct a novel benchmark to facilitate the evaluation of generalizable RL algorithms in practical building control tasks. Our results show that existing multi-objective RL methods are capable of achieving reasonable trade-offs between conflicting objectives. However, their performance degrades under certain environment variations, underscoring the importance of incorporating dynamics-dependent contextual information into the policy learning process.

📄 PDF Abstract BibTeX arXiv:2507.07769

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Web-imageability of the Behavioral Features of Basic-level Concepts

2014-05-01 · LREC 2014 5 · Yoshihiko Hayashi

The recent research direction toward multimodal semantic representation would be further advanced, if we could have a machinery to collect adequate images from the Web, given a target concept. With this motivation, this …

FormInformation Retrieval

BEAVER: A Training-Free Hierarchical Prompt Compression Method via Structure-Aware Page Selection

2026-03-20 · Zhengpei Hu, Kai Li, Dapeng Fu, Chang Zeng 외 arxiv

The exponential expansion of context windows in LLMs has unlocked capabilities for long-document understanding but introduced severe bottlenecks in inference latency and information utilization. Existing compression meth…

BEAVER: An Efficient Deterministic LLM Verifier

2025-12-05 · Tarun Suresh, Nalin Wadhwa, Debangshu Banerjee, Gagandeep Singh arxiv

As large language models (LLMs) transition from research prototypes to production systems, practitioners often need reliable methods to verify model outputs and characterize tail risk for safe deployment. While sampling-…

BEAVER: An Enterprise Benchmark for Text-to-SQL

2024-09-03 · Peter Baile Chen, Fabian Wenz, Yi Zhang, Devin Yang 외

Existing text-to-SQL benchmarks have largely been constructed from web tables with human-generated question-SQL pairs. LLMs typically show strong results on these benchmarks, leading to a belief that LLMs are effective a…

Natural Language QueriesPrompt EngineeringRAGText to SQL+1

Building Agent Harnesses for Scientific Curation from Multimodal Sources

2026-06-19 · Sheng Zhang, Qin Liu, Renqian Luo, Shufang Xie 외 arxiv

Scientific discovery workflows often depend on structured curation from the literature. This is difficult for current agents because the key evidence is scattered across long text, dense tables, and figures, and the fina…