Countable state Markov decision processes with unbounded jump rates and discounted cost: optimality equation and approximations-Reference-Cited by-同舟云学术

Countable state Markov decision processes with unbounded jump rates and discounted cost: optimality equation and approximations

Published:2015-12 Issue:04 Volume:47 Page:1088-1107
ISSN:0001-8678
Container-title:Advances in Applied Probability
language:en
Short-container-title:Adv. Appl. Probab.

Author:

Blok H.,Spieksma F. M.

Abstract

This paper considers Markov decision processes (MDPs) with unbounded rates, as a function of state. We are especially interested in studying structural properties of optimal policies and the value function. A common method to derive such properties is by value iteration applied to the uniformised MDP. However, due to the unboundedness of the rates, uniformisation is not possible, and so value iteration cannot be applied in the way we need. To circumvent this, one can perturb the MDP. Then we need two results for the perturbed sequence of MDPs: 1. there exists a unique solution to the discounted cost optimality equation for each perturbation as well as for the original MDP; 2. if the perturbed sequence of MDPs converges in a suitable manner then the associated optimal policies and the value function should converge as well. We can model both the MDP and perturbed MDPs as a collection of parametrised Markov processes. Then both of the results above are essentially implied by certain continuity properties of the process as a function of the parameter. In this paper we deduce tight verifiable conditions that imply the necessary continuity properties. The most important of these conditions are drift conditions that are strongly related to nonexplosiveness.

Publisher

Cambridge University Press (CUP)

Subject

Applied Mathematics,Statistics and Probability

Reference16 articles.

1. Discounted continuous-time Markov decision processes with unbounded rates and randomized history-dependent policies: the dynamic programming approach

2. A survey of recent results on continuous-time Markov decision processes

3. Denumerable-state continuous-time Markov decision processes with unbounded transition and reward rates under the discounted criterion

4. Discounted Continuous-Time Markov Decision Processes with Constraints: Unbounded Transition and Loss Rates

Cited by 9 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Scheduling servers in a two-stage queue with abandonments and costs;Probability in the Engineering and Informational Sciences;2022-07-20

2. Optimal speed profile of a DVFS processor under soft deadlines;Performance Evaluation;2021-12

3. Optimal control of admission in service in a queue with impatience and setup costs;Performance Evaluation;2020-12

4. Optimal Control of Parallel Queues for Managing Volunteer Convergence;Production and Operations Management;2020-07-02

5. K competing queues with customer abandonment: optimality of a generalised $$c \mu $$ c μ -rule by the Smoothed Rate Truncation method;Annals of Operations Research;2019-01-11