paper-with-me

홈 › Papers

Optimal Policy Synthesis from A Sequence of Goal Sets with An Application to Electric Distribution System Restoration

2024-04-05 · İlker Işık, Onur Yigit Arpali, Ebru Aydin Gol

Motivated by the post-disaster distribution system restoration problem, in this paper, we study the problem of synthesizing the optimal policy for a Markov Decision Process (MDP) from a sequence of goal sets. For each goal set, our aim is to both maximize the probability to reach and minimize the expected time to reach the goal set. The order of the goal sets represents their priority. In particular, our aim is to generate a policy that is optimal with respect to the first goal set, and it is optimal with respect to the second goal set among the policies that are optimal with respect to the first goal set and so on. To synthesize such a policy, we iteratively filter the applicable actions according to the goal sets. We illustrate the developed method over sample distribution systems and disaster scenarios.

📄 PDF Abstract BibTeX arXiv:2404.04338

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

In Search of Trees: Decision-Tree Policy Synthesis for Black-Box Systems via Search

2024-09-05 · Emir Demirović, Christian Schilling, Anna Lukina

Decision trees, owing to their interpretability, are attractive as control policies for (dynamical) systems. Unfortunately, constructing, or synthesising, such policies is a challenging task. Previous approaches do so by…

Transfer Reinforcement Learning in Heterogeneous Action Spaces using Subgoal Mapping

2024-10-18 · Kavinayan P. Sivakumar, Yan Zhang, Zachary Bell, Scott Nivison 외

In this paper, we consider a transfer reinforcement learning problem involving agents with different action spaces. Specifically, for any new unseen task, the goal is to use a successful demonstration of this task by an …

reinforcement-learningReinforcement LearningTransfer LearningTransfer Reinforcement Learning

Synthesizing Action Sequences for Modifying Model Decisions

2019-09-30 · Goutham Ramakrishnan, Yun Chan Lee, Aws Albarghouthi

When a model makes a consequential decision, e.g., denying someone a loan, it needs to additionally generate actionable, realistic feedback on what the person can do to favorably change the decision. We cast this problem…

modelProgram Synthesis

Reinforcement Learning of Control Policy for Linear Temporal Logic Specifications Using Limit-Deterministic Generalized Büchi Automata

2020-01-14 · Ryohei Oura, Ami Sakakibara, Toshimitsu Ushio

This letter proposes a novel reinforcement learning method for the synthesis of a control policy satisfying a control specification described by a linear temporal logic formula. We assume that the controlled system is mo…

Reinforcement Learning

Planning as Descent: Goal-Conditioned Latent Trajectory Synthesis in Learned Energy Landscapes

2025-12-19 · Carlos Vélez García, Miguel Cazorla, Jorge Pomares arxiv

We present Planning as Descent (PaD), a framework for offline goal-conditioned reinforcement learning that grounds trajectory synthesis in verification. Instead of learning a policy or explicit planner, PaD learns a goal…

Reinforcement Learning