3A review of stochastic algorithms with continuous value function approximation and some new approximate policy iteration algorithms for multidimensional continuous applications336-352
5Online optimal control of nonlinear discrete-time systems using approximate dynamic programming361-369
6Approximate dynamic programming solutions with a single network adaptive critic for a class of nonlinear systems370-380
7Finite horizon optimal control of discrete-time nonlinear systems with unfixed initial state using adaptive dynamic programming381-390
8A model-based approximate λ-policy iteration approach to online evasive path planning and the video game Ms.Pac-Man391-399
12Multiresolution state-space discretization for Q-Learning with pseudorandomized discretization431-439
13Hierarchical state-abstracted and socially augmented Q-Learning for reducing complexity in agent-based learning440-450
14Moving least-squares approximations for linearly-solvable stochastic optimal control problems451-463