paper-with-me

홈 › Papers

Optimizing Task Completion Time Updates Using POMDPs

2026-03-12 · Duncan Eddy, Esen Yel, Emma Passmore, Niles Egan, Grayson Armour, Dylan M. Asmar, Mykel J. Kochenderfer arxiv

Managing announced task completion times is a fundamental control problem in project management. While extensive research exists on estimating task durations and task scheduling, the problem of when and how to update completion times communicated to stakeholders remains understudied. Organizations must balance announcement accuracy against the costs of frequent timeline updates, which can erode stakeholder trust and trigger costly replanning. Despite the prevalence of this problem, current approaches rely on static predictions or ad-hoc policies that fail to account for the sequential nature of announcement management. In this paper, we formulate the task announcement problem as a Partially Observable Markov Decision Process (POMDP) where the control policy must decide when to update announced completion times based on noisy observations of true task completion. Since most state variables (current time and previous announcements) are fully observable, we leverage the Mixed Observability MDP (MOMDP) framework to enable more efficient policy optimization. Our reward structure captures the dual costs of announcement errors and update frequency, enabling synthesis of optimal announcement control policies. Using off-the-shelf solvers, we generate policies that act as feedback controllers, adaptively managing announcements based on belief state evolution. Simulation results demonstrate significant improvements in both accuracy and announcement stability compared to baseline strategies, achieving up to 75\% reduction in unnecessary updates while maintaining or improving prediction accuracy.

📄 PDF Abstract BibTeX arXiv:2603.12340

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Engagement Process: Rethinking the Temporal Interface of Action and Observation

2026-05-12 · Jialian Li, Yuchen Cao, Junhong Liu, Weiran Guo 외 arxiv

Task completion in digital and physical environments increasingly involves complex temporal interaction, where actions and observations unfold over different time scales rather than align with fixed observation--action s…

Recursively-Constrained Partially Observable Markov Decision Processes

2023-10-15 · Qi Heng Ho, Tyler Becker, Benjamin Kraske, Zakariya Laouar 외

Many sequential decision problems involve optimizing one objective function while imposing constraints on other objectives. Constrained Partially Observable Markov Decision Processes (C-POMDP) model this case with transi…

Technical Report: The Policy Graph Improvement Algorithm

2020-09-04 · Joni Pajarinen

Optimizing a partially observable Markov decision process (POMDP) policy is challenging. The policy graph improvement (PGI) algorithm for POMDPs represents the policy as a fixed size policy graph and improves the policy …

Monte-Carlo Planning in Large POMDPs

2010-12-01 · NeurIPS 2010 12 · David Silver, Joel Veness

This paper introduces a Monte-Carlo algorithm for online planning in large POMDPs. The algorithm combines a Monte-Carlo update of the agent's belief state with a Monte-Carlo tree search from the current belief state. The…

Optimizing Sequential Medical Treatments with Auto-Encoding Heuristic Search in POMDPs

2019-05-17 · Luchen Li, Matthieu Komorowski, Aldo A. Faisal

Health-related data is noisy and stochastic in implying the true physiological states of patients, limiting information contained in single-moment observations for sequential clinical decision making. We model patient-cl…

Decision MakingHeuristic Search