A review of stochastic algorithms with continuous value function approximation and some new approximate policy iteration algorithms for multidimensional continuous applications336-352
Online optimal control of nonlinear discrete-time systems using approximate dynamic programming361-369
Approximate dynamic programming solutions with a single network adaptive critic for a class of nonlinear systems370-380
Finite horizon optimal control of discrete-time nonlinear systems with unfixed initial state using adaptive dynamic programming381-390
A model-based approximate λ-policy iteration approach to online evasive path planning and the video game Ms.Pac-Man391-399