paper-with-me

Papers

Efficient Solving of Large Single Input Superstate Decomposable Markovian Decision Process

2025-08-01 · Youssef Ait El Mahjoub, Jean-Michel Fourneau, Salma Alouah arxiv

Solving Markov Decision Processes (MDPs) remains a central challenge in sequential decision-making, especially when dealing with large state spaces and long-term optimization criteria. A key step in Bellman dynamic programming algorithms is the policy evaluation, which becomes computationally demanding in infinite-horizon settings such as average-reward or discounted-reward formulations. In the context of Markov chains, aggregation and disaggregation techniques have for a long time been used to reduce complexity by exploiting structural decompositions. In this work, we extend these principles to a structured class of MDPs. We define the Single-Input Superstate Decomposable Markov Decision Process (SISDMDP), which combines Chiu's single-input decomposition with Robertazzi's single-cycle recurrence property. When a policy induces this structure, the resulting transition graph can be decomposed into interacting components with centralized recurrence. We develop an exact and efficient policy evaluation method based on this structure. This yields a scalable solution applicable to both average and discounted reward MDPs.

📄 PDF Abstract BibTeX arXiv:2508.00816

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Model-Based Learning of Near-Optimal Finite-Window Policies in POMDPs

2026-04-01 · Philip Jordan, Maryam Kamgarpour arxiv

We study model-based learning of finite-window policies in tabular partially observable Markov decision processes (POMDPs). A common approach to learning under partial observability is to approximate unbounded history de…

Scalable Policy-Based RL Algorithms for POMDPs

2025-10-08 · Ameya Anjarlekar, Rasoul Etesami, R Srikant arxiv

The continuous nature of belief states in POMDPs presents significant computational challenges in learning the optimal policy. In this paper, we consider an approach that solves a Partially Observable Reinforcement Learn…

Reinforcement Learning

Efficient Encoder-Decoder Transformer Decoding for Decomposable Tasks

2024-03-19 · Bo-Ru Lu, Nikita Haduong, Chien-Yu Lin, Hao Cheng 외

Transformer-based NLP models are powerful but have high computational costs that limit deployment. Finetuned encoder-decoder models are popular in specialized domains and can outperform larger more generalized decoder-on…

DecoderDialogue State TrackingQuestion Answering

Point-Query Quadtree for Crowd Counting, Localization, and More

2023-08-26 · ICCV 2023 1 · Chengxin Liu, Hao Lu, Zhiguo Cao, Tongliang Liu

We show that crowd counting can be viewed as a decomposable point querying process. This formulation enables arbitrary points as input and jointly reasons whether the points are crowd and where they locate. The querying …

Crowd Counting

Decomposable Submodular Maximization in Federated Setting

2024-01-31 · Akbar Rafiey

Submodular functions, as well as the sub-class of decomposable submodular functions, and their optimization appear in a wide range of applications in machine learning, recommendation systems, and welfare maximization. Ho…

Recommendation Systems