Another set of conditions for strong n (n =-1, 0) discount optimality in Markov decision processes

Zhu, QX; Guo, XP<sup>*</sup>

doi:10.1080/07362990500184865

摘要

In this paper we study discrete-time Markov decision processes with average expected costs (AEC) and discount-sensitive criteria in Borel state and action spaces. The costs may have neither upper nor lower bounds. We propose another set of conditions on the system's primitive data, and under which we prove (1) AEC optimality and strong -1-discount optimality are equivalent; (2) a condition equivalent to strong 0-discount optimal stationary policies; and (3) the existence of strong n (n = -1, 0)-discount optimal stationary policies. Our conditions are weaker than those in the previous literature. In particular, the "stochastic monotonicity condition" in this paper has been first used to study strong n (n = -1, 0)-discount optimality. Moreover, we provide a new approach to prove the existence of strong 0-discount optimal stationary policies. It should be noted that our way is slightly different from those in the previous literature. Finally, we apply our results to an inventory system and a controlled queueing system.

出版日期2005
单位中山大学

全文

访问全文

收藏分享被引(12) 浏览

更新时间：2019-09-03 21:03

Another set of conditions for strong n (n =-1, 0) discount optimality in Markov decision processes

摘要

全文

产品服务

站内浏览

服务支持

联系方式

科研之友