The value function for the average cost control of a class of partially observed Markov chains is derived as the "vanishing discount limit," in a suitable sense, of the value functions for the corresponding discounted cost problems. The limiting procedure is justified by bounds derived using a simple coupling argument.
No takes yet. Share an insight, caveat, or question.
Vivek S. Borkar (2000) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: