The maximum principle (MP) for the discrete-time stochastic optimal control problems is proved. It is shown that the adjoint equations of the MP are a pair of backward stochastic difference equations.
No takes yet. Share an insight, caveat, or question.
Lin et al. (2014) studied this question.
Synapse has enriched one closely related paper. Consider it for comparative context: