Semi-Markov Decision Processes

Book Chapter

Semi-Markov Decision Processes

Springer US, (2007), 105-120

DOI: 10.1007/978-0-387-36951-8_5

N/ACitations

2Readers

Get full text

Abstract

The underlying stochastic processes in DTMDPs are discrete time Markov chains, where the decision epochs are equally periodic or the length of adjacent decision epochs are not considered. Those in CTMDPs are continuous time Markov chains, where the decision is chosen every time. In this chapter, we study a stationary semi-Markov decision processes (SMDPs) model, where the underlying stochastic processes are semi-Markov processes. Here, the decision epoch is exactly the state transition epoch with its length being random. We transform the SMDP model into a stationary DTMDP model for either the total reward criterion or the average criterion, similarly to the stationary CTMDP model with the average criterion discussed in Section 4.3. Thus, the results in DTMDP can be used directly for SMDP for the discounted criterion, the total reward criterion, and the average criterion.

Cite

CITATION STYLE

APA

Semi-Markov Decision Processes. (2007). In Markov Decision Processes With Their Applications (pp. 105–120). Springer US. https://doi.org/10.1007/978-0-387-36951-8_5

Semi-Markov Decision Processes

Abstract

Cite

Register to see more suggestions