paper-with-me

홈 › Papers

Safe Learning for Near Optimal Scheduling

2020-05-19 · Damien Busatto-Gaston, Debraj Chakraborty, Shibashis Guha, Guillermo A. Pérez, Jean-François Raskin

In this paper, we investigate the combination of synthesis, model-based learning, and online sampling techniques to obtain safe and near-optimal schedulers for a preemptible task scheduling problem. Our algorithms can handle Markov decision processes (MDPs) that have 1020 states and beyond which cannot be handled with state-of-the art probabilistic model-checkers. We provide probably approximately correct (PAC) guarantees for learning the model. Additionally, we extend Monte-Carlo tree search with advice, computed using safety games or obtained using the earliest-deadline-first scheduler, to safely explore the learned model online. Finally, we implemented and compared our algorithms empirically against shielded deep Q-learning on large task systems.

📄 PDF Abstract BibTeX arXiv:2005.09253

Code (0)

등록된 구현이 없습니다.

Tasks

Q-LearningScheduling

Similar Papers 제목 키워드 기반

Scheduling for Urban Air Mobility using Safe Learning

2022-09-28 · Surya Murthy, Natasha A. Neogi, Suda Bharadwaj

This work considers the scheduling problem for Urban Air Mobility (UAM) vehicles travelling between origin-destination pairs with both hard and soft trip deadlines. Each route is described by a discrete probability distr…

Scheduling

Safety Verification and Control for Collision Avoidance at Road Intersections

2016-12-08 · Heejin Ahn, Domitilla Del Vecchio

This paper presents the design of a supervisory algorithm that monitors safety at road intersections and overrides drivers with a safe input when necessary. The design of the supervisor consists of two parts: safety veri…

BlockingCollision AvoidanceScheduling

Data Driven Safe Gain-Scheduling Control

2022-01-03 · Amir Modares, Nasser Sadati, Hamidreza Modares

Data-based safe gain-scheduling controllers are presented for discrete-time linear parameter-varying systems (LPV) with polytopic models. First, $\lambda$-contractivity conditions are provided under which safety and stab…

Scheduling

Separation is Optimal for LQR under Intermittent Feedback

2026-03-29 · Abdullah Y. Etcibasi, C. Emre Koksal, Eylem Ekici arxiv

We study finite-horizon linear-quadratic regulation of a scalar linear system with intermittent state feedback under an average communication-rate constraint. In this setting, the scheduling policy and controller are gen…

Bi-Level Online Provisioning and Scheduling with Switching Costs and Cross-Level Constraints

2026-01-26 · Jialei Liu, C. Emre Koksal, Ming Shi arxiv

We study a bi-level online provisioning and scheduling problem motivated by network resource allocation, where provisioning decisions are made at a slow time scale while queue-/state-dependent scheduling is performed at …