Uniqueness and Stability of Optimal Policies of Finite State Markov Decision Processes
摘要
In this chapter we study infinite horizon discrete-time optimal control of Markov Decision Processes (MDPs) with finite state spaces and compact action sets and employ the long-run expected average cost criterion.