Continuous-time Markov decision processes with discounted rewards: The case of Polish spaces

被引：37

作者：

Guo, Xianping ^{[1
]}

机构：

[1] Zhongshan Univ, Sch Math & Computat Sci, Guangzhou 510275, Peoples R China

来源：

MATHEMATICS OF OPERATIONS RESEARCH | 2007年 / 32卷 / 01期

关键词：

Q-process; general state space; discounted reward; optimal stationary policy;

D O I：

10.1287/moor.1060.0210

中图分类号：

C93 [管理学]; O22 [运筹学];

学科分类号：

070105 ; 12 ; 1201 ; 1202 ; 120202 ;

摘要：

This paper deals with continuous-time Markov decision processes in Polish spaces, under an expected discounted reward criterion. The transition rates of underlying continuous-time jump Markov processes are allowed to be unbounded, and the reward rates may have neither upper nor lower bounds. We first give conditions on the controlled system's primitive data. Under these conditions we prove that the transition functions of possibly nonhomogeneous continuous-time Markov processes are regular by using Feller's construction approach to such transition functions. Then, under additional continuity and compactness conditions, we ensure the existence of optimal stationary policies by using the technique of extended infinitesimal operators associated with the transition functions, and also provide a recursive way to compute (or at least to approximate) the optimal reward values. Finally, we use examples to illustrate our results and the gap between our conditions and those in the previous literature.

引用

页码：73 / 87

页数：15

共 50 条

[1] Discounted optimality for continuous-time Markov decision processes in Polish spaces
Guo, Xianping
2006 CHINESE CONTROL CONFERENCE, VOLS 1-5, 2006, : 1785 - 1787
[2] DISCOUNTED CONTINUOUS-TIME CONSTRAINED MARKOV DECISION PROCESSES IN POLISH SPACES
Guo, Xianping
Song, Xinyuan
ANNALS OF APPLIED PROBABILITY, 2011, 21 (05): : 2016 - 2049
[3] Continuous-time fuzzy decision processes with discounted rewards
Yoshida, Y
FUZZY SETS AND SYSTEMS, 2003, 139 (02) : 333 - 348
[4] Average optimality for continuous-time Markov decision processes in Polish spaces
Guo, Xianping
Rieder, Ulrich
ANNALS OF APPLIED PROBABILITY, 2006, 16 (02): : 730 - 756
[5] Continuous time Markov decision processes with expected discounted total rewards
Hu, QY
Liu, JY
Yue, WY
COMPUTATIONAL SCIENCE - ICCS 2003, PT II, PROCEEDINGS, 2003, 2658 : 64 - 73
[6] Average optimality inequality for continuous-time Markov decision processes in Polish spaces
Quanxin Zhu
Mathematical Methods of Operations Research, 2007, 66 : 299 - 313
[7] Average optimality inequality for continuous-time Markov decision processes in Polish spaces
Zhu, Quanxin
MATHEMATICAL METHODS OF OPERATIONS RESEARCH, 2007, 66 (02) : 299 - 313
[8] Continuous-time controlled Markov chains with discounted rewards
Guo, XP
Hernández-Lerma, O
ACTA APPLICANDAE MATHEMATICAE, 2003, 79 (03) : 195 - 216
[9] Continuous-Time Controlled Markov Chains with Discounted Rewards
Xianping Guo
Onésimo Hernández-Lerma
Acta Applicandae Mathematica, 2003, 79 : 195 - 216
[10] Zero-sum games for continuous-time jump Markov processes in polish spaces: Discounted payoffs
Guo, Xianping
Hernandez-Lerma, Onesimo
ADVANCES IN APPLIED PROBABILITY, 2007, 39 (03) : 645 - 668

← 1 2 3 4 5 →