A review of stochastic algorithms with continuous value function approximation and some new approximate policy iteration algorithms for multidimensional continuous applications
Warren B. POWELL
Jun MA
Keywords:Approximate dynamic programmingReinforcement learningOptimal controlApproximation algorithms
Publication Date:2011-01-01
Online Publishing Date:2025-08-15(First online date of this platform, not the publication date of the document)
Pages:17( 336-352 )
