This paper proposes bounds and action elimination procedures for policy iteration and modified policy iteration. Procedures to eliminate nonoptimal actions for one iteration and for all subsequent iterations are presented. The implementation of these procedures is discussed and encouraging computational results are presented.
No takes yet. Share an insight, caveat, or question.
Puterman et al. (1982) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: