Average optimality inequality for continuous-time Markov decision processes in Polish spaces

被引：12

作者：

Zhu, Quanxin ^{[1
]}

机构：

[1] S China Normal Univ, Dept Math, Guangzhou 510631, Peoples R China

来源：

MATHEMATICAL METHODS OF OPERATIONS RESEARCH | 2007年 / 66卷 / 02期

关键词：

continuous-time Markov decision process; average optimality inequality; general state space; unbounded cost; optimal stationary policy;

D O I：

10.1007/s00186-007-0157-x

中图分类号：

C93 [管理学]; O22 [运筹学];

学科分类号：

070105 ; 12 ; 1201 ; 1202 ; 120202 ;

摘要：

In this paper, we study the average optimality for continuous-time controlled jump Markov processes in general state and action spaces. The criterion to be minimized is the average expected costs. Both the transition rates and the cost rates are allowed to be unbounded. We propose another set of conditions under which we first establish one average optimality inequality by using the well-known "vanishing discounting factor approach". Then, when the cost (or reward) rates are nonnegative (or nonpositive), from the average optimality inequality we prove the existence of an average optimal stationary policy in all randomized history dependent policies by using the Dynkin formula and the Tauberian theorem. Finally, when the cost (or reward) rates have neither upper nor lower bounds, we also prove the existence of an average optimal policy in all (deterministic) stationary policies by constructing a "new" cost (or reward) rate.

引用

页码：299 / 313

页数：15

共 50 条

[1] Average optimality inequality for continuous-time Markov decision processes in Polish spaces
Quanxin Zhu
Mathematical Methods of Operations Research, 2007, 66 : 299 - 313
[2] Average optimality for continuous-time Markov decision processes in Polish spaces
Guo, Xianping
Rieder, Ulrich
ANNALS OF APPLIED PROBABILITY, 2006, 16 (02): : 730 - 756
[3] Average sample-path optimality for continuous-time Markov decision processes in Polish spaces
Zhu, Quan-xin
ACTA MATHEMATICAE APPLICATAE SINICA-ENGLISH SERIES, 2011, 27 (04): : 613 - 624
[4] Average sample-path optimality for continuous-time Markov decision processes in Polish spaces
Quan-xin Zhu
Acta Mathematicae Applicatae Sinica, English Series, 2011, 27 : 613 - 624
[5] Discounted optimality for continuous-time Markov decision processes in Polish spaces
Guo, Xianping
2006 CHINESE CONTROL CONFERENCE, VOLS 1-5, 2006, : 1785 - 1787
[6] Bias and overtaking optimality for continuous-time jump Markov decision processes in polish spaces
Zhu, Quanxin
Prieto-Rumeau, Tomas
JOURNAL OF APPLIED PROBABILITY, 2008, 45 (02) : 417 - 429
[7] Policy Iteration for Continuous-Time Average Reward Markov Decision Processes in Polish Spaces
Zhu, Quanxin
Yang, Xinsong
Huang, Chuangxia
ABSTRACT AND APPLIED ANALYSIS, 2009,
[8] Verifiable conditions for average optimality of continuous-time Markov decision processes
Zou, Xiaolong
Huang, Yonghui
OPERATIONS RESEARCH LETTERS, 2016, 44 (06) : 742 - 746
[9] STRONG AVERAGE OPTIMALITY CRITERION FOR CONTINUOUS-TIME MARKOV DECISION PROCESSES
Wei, Qingda
Chen, Xian
KYBERNETIKA, 2014, 50 (06) : 950 - 977
[10] New sufficient conditions for average optimality in continuous-time Markov decision processes
Ye, Liuer
Guo, Xianping
MATHEMATICAL METHODS OF OPERATIONS RESEARCH, 2010, 72 (01) : 75 - 94

← 1 2 3 4 5 →