Continuous adaptive critic designs
Hanselmann, T., Noakes, L. and Zaknich, A. (2005) Continuous adaptive critic designs. In: International Joint Conference on Neural Networks, IJCNN 2005, 31 July - 4 August, Montreal, Canada.
|PDF - Published Version |
Download (1397kB) | Preview
*Subscription may be required
A continuous formulation of an adaptive critic design (ACD) is investigated. Connections to the discrete case are made, where backpropagation through time (BPTT) and realtime recurrent learning (RTRL) are prevalent. A second order actor adaptation, based on Newton's method, is established for fast actor convergence. Also, a fast critic update for concurrent actor-critic training is outlined that keeps the Bellman optimality correct to first order approximation after actor changes.
|Publication Type:||Conference Paper|
|Murdoch Affiliation:||School of Engineering|
|Copyright:||© 2005 IEEE|
|Notes:||In Proceedings of the International Joint Conference on Neural Networks, 2005. IJCNN '05, Pages 3001-3006.|
|Item Control Page|