Adaptive dynamic programming: an introduction

Fei-Yue Wang,Huaguang Zhang,Derong Liu
DOI: https://doi.org/10.1109/MCI.2009.932261
IF: 9.809
2009-01-01
IEEE Computational Intelligence Magazine
Abstract:In this article, we introduce some recent research trends within the field of adaptive/approximate dynamic programming (ADP), including the variations on the structure of ADP schemes, the development of ADP algorithms and applications of ADP schemes. For ADP algorithms, the point of focus is that iterative algorithms of ADP can be sorted into two classes: one class is the iterative algorithm with initial stable policy; the other is the one without the requirement of initial stable policy. It is generally believed that the latter one has less computation at the cost of missing the guarantee of system stability during iteration process. In addition, many recent papers have provided convergence analysis associated with the algorithms developed. Furthermore, we point out some topics for future studies.
What problem does this paper attempt to address?