Hostname: page-component-745bb68f8f-v2bm5 Total loading time: 0 Render date: 2025-01-08T05:02:50.033Z Has data issue: false hasContentIssue false

Average cost semi-markov decision processes

Published online by Cambridge University Press:  14 July 2016

Sheldon M. Ross*
Affiliation:
University of California, Berkeley

Abstract

The semi-Markov decision model is considered under the criterion of long-run average cost. A new criterion, which for any policy considers the limit of the expected cost incurred during the first n transitions divided by the expected length of the first n transitions, is considered. Conditions guaranteeing that an optimal stationary (non-randomized) policy exist are then presented. It is also shown that the above criterion is equivalent to the usual one under certain conditions.

Type
Research Papers
Copyright
Copyright © Applied Probability Trust 1970 

Access options

Get access to the full version of this content by using one of the access options below. (Log in options will check for institutional or personal access. Content may require purchase if you do not have access.)

References

[1] Blackwell, D. (1965) Discounted dynamic programming. Ann. Math. Statist. 36, 226235.Google Scholar
[2] Derman, C. (1966) Denumerable state Markovian decision processes–average cost criterion. Ann. Math. Statist. 37, 15451554.Google Scholar
[3] Howard, R. (1963) Semi-Markovian decision processes. Bull. Inst. Internat. Statist. 40, 625652.Google Scholar
[4] Jewell, W. S. (1963) Markov renewal programming I and II. Operat. Res. 2, 938971.Google Scholar
[5] Ross, S. M. (1968) Non-discounted denumerable Markovian decision models. Ann. Math. Statist. 39, 412423.CrossRefGoogle Scholar
[6] Ross, S. M. (1968) Arbitrary state Markovian decision processes. Ann. Math. Statist. 39, 21182122.Google Scholar