Value iteration in average cost Markov control processes on Borel spaces期刊界 All Journals 搜尽天下杂志传播学术成果专业期刊搜索期刊信息化学术搜索

按检索

Value iteration in average cost Markov control processes on Borel spaces

Authors:	Raúl Montes-de-Oca Onésimo Hernández-Lerma

Institution:	(1) Departamento de Matemáticas, Universidad Autónoma Metropolitana-Iztapalapa, A.Postal 55534, 09340 México, D.F., Mexico;(2) Departamento de Matemáticas, CINVESTAV-IPN, A.Postal 14-740, 07000 México, D.F., Mexico

Abstract:	This paper deals with discrete-time Markov control processes withBorel state and control spaces, with possiblyunbounded costs andnoncompact control constraint sets, and the average cost criterion. Conditions are given for the convergence of the value iteration algorithm to the optimal average cost, and for a sequence of finite-horizon optimal policies to have an accumulation point which is average cost optimal.This research was partially supported by the Consejo Nacional de Ciencia y Tecnología (CONACyT) under grant 1332-E9206.

Keywords:	93E20 90C40
本文献已被 SpringerLink 等数据库收录！