paper-with-me

Papers

Solving the non-preemptive two queue polling model with generally distributed service and switch-over durations and Poisson arrivals as a Semi-Markov Decision Process

2021-12-13 · Dylan Solms

The polling system with switch-over durations is a useful model with several practical applications. It is classified as a Discrete Event Dynamic System (DEDS) for which no one agreed upon modelling approach exists. Furthermore, DEDS are quite complex. To date, the most sophisticated approach to modelling the polling system of interest has been a Continuous-time Markov Decision Process (CTMDP). This paper presents a Semi-Markov Decision Process (SMDP) formulation of the polling system as to introduce additional modelling power. Such power comes at the expense of truncation errors and expensive numerical integrals which naturally leads to the question of whether the SMDP policy provides a worthwhile advantage. To further add to this scenario, it is shown how sparsity can be exploited in the CTMDP to develop a computationally efficient model. The discounted performance of the SMDP and CTMDP policies are evaluated using a Semi-Markov Process simulator. The two policies are accompanied by a heuristic policy specifically developed for this polling system a well as an exhaustive service policy. Parametric and non-parametric hypothesis tests are used to test whether differences in performance are statistically significant.

📄 PDF Abstract BibTeX arXiv:2112.06578

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Golden Queue Managers 설명 없음

Similar Papers 제목 키워드 기반

Optimization of Traffic Control in MMAP[c]/PH[c]/S Catastrophic Queueing Model with PH Retrial Times and Controllable Preemptive Repeat Priority Policy

2022-03-04 · Raina Raj, Vidyottama Jain

The presented study elaborates a multi-server catastrophic retrial queueing model considering preemptive repeat priority policy with phase-type (PH) distributed retrial times. For the sake of comprehension, the scenario …

Fast Distributed Inference Serving for Large Language Models

2023-05-10 · Bingyang Wu, Yinmin Zhong, Zili Zhang, Shengyu Liu 외

Large language models (LLMs) power a new generation of interactive AI applications exemplified by ChatGPT. The interactive nature of these applications demands low latency for LLM inference. Existing LLM serving systems …

BlockingGPUManagementScheduling

A Queueing Model for the Ambulance Ramping Problem with an Offload Zone

2024-01-12 · Josef Zuk, David Kirszenblat

This work develops a methodology for studying the effect of an offload zone on the ambulance ramping problem using a multi-server, multi-class non-preemptive priority queueing model that can be treated analytically. A pr…

Mining Voter Behaviour and Confidence: A Rule-Based Analysis of the 2022 U.S. Elections

2025-07-17 · Md Al Jubair, Mohammad Shamsul Arefin, Ahmed Wasif Reza arxiv

This study explores the relationship between voter trust and their experiences during elections by applying a rule-based data mining technique to the 2022 Survey of the Performance of American Elections (SPAE). Using the…

PecSched: Preemptive and Efficient Cluster Scheduling for LLM Inference

2024-09-23 · Zeyu Zhang, Haiying Shen

The scaling of transformer-based Large Language Models (LLMs) has significantly expanded their context lengths, enabling applications where inputs exceed 100K tokens. Our analysis of a recent Azure LLM inference trace re…

2kBlockingLarge Language ModelScheduling