Abstract:In this paper, a novel adaptive dynamic programming (ADP)-based optimal control method is developed for discrete-time systems subject to constraints and disturbances. Particularly, a safe policy iteration scheme is designed to handle state and input constraints, including both hard and soft constraints, by converting the original policy improvement strategy into a constrained optimization problem with a prescribed state cost function. After that, an actor-critic-disturbance framework is introduced to address the constrained optimal control problem. The robust safety against disturbances is treated as a two-player zero-sum game, where the actor and disturbance neural networks are used to approximate the optimal control input and the disturbance policy, respectively. The convergence property of the proposed algorithm is analyzed, and the multi-step version of the proposed ADP scheme is derived based on this property. Simulation results are demonstrated and discussed to validate the effectiveness and performance of the proposed method. Note to Practitioners—Addressing constraints in optimal control problems is essential for guaranteeing the safe operation of controlled systems. However, conventional ADP algorithms struggle to simultaneously manage state and control input constraints during the search for the optimal solution. In real-world applications, another critical and common issue is the presence of external disturbances, where disturbances that cause the control object to deviate from the safe region must be constrained while seeking an optimal control policy. Bearing these factors in mind, this study presents a novel ADP scheme for solving optimal control problems of discrete-time systems, taking into account state and control constraints as well as the impact of disturbances. Moreover, the convergence analysis of the proposed SADP scheme is provided, offering a powerful theoretical foundation for guaranteeing the safety and feasibility of the controlled system during operation.

Discrete‐Time Optimal Control of State‐Constrained Nonlinear Systems Using Approximate Dynamic Programming

Optimal Control for Constrained Discrete-Time Nonlinear Systems Based on Safe Reinforcement Learning.

Adaptive Dynamic Programming for Nonaffine Nonlinear Optimal Control Problem with State Constraints

Adaptive dynamic programming for optimal control of discrete‐time nonlinear system with state constraints based on control barrier function

Adaptive Finite-Time Optimised Impedance Control for Robotic Manipulators with State Constraints

ADP-Based Optimal Control for Discrete-Time Systems With Safe Constraints and Disturbances

Model-free Adaptive Dynamic Programming for Optimal Control of Discrete-time Affine Nonlinear System

Approximate Dynamic Programming for Constrained Piecewise Affine Systems with Stability and Safety Guarantees

Intelligent Optimal Control of Constrained Nonlinear Systems Via Receding-Horizon Heuristic Dynamic Programming

Output Constrained Adaptive Dynamic Programming For Continuous-Time Nonlinear Systems

A Novel Approximate Dynamic Programming Structure for Optimal Control of Discrete-Time Time-Varying Nonlinear Systems

A New Approach to Finite-Horizon Optimal Control for Discrete-Time Affine Nonlinear Systems via a Pseudolinear Method

Event-triggered design for discrete-time nonlinear systems with control constraints

Costate-Supplement ADP for Model-Free Optimal Control of Discrete-Time Nonlinear Systems

Adaptive dynamic programming-based optimal control for nonlinear state constrained systems with input delay

Self-Triggered Approximate Optimal Neuro-Control for Nonlinear Systems Through Adaptive Dynamic Programming

Policy-Iteration-Based Finite-Horizon Approximate Dynamic Programming for Continuous-Time Nonlinear Optimal Control

A hybrid model-based optimal control method for nonlinear systems using simultaneous dynamic optimization strategies

Robust ADP Design for Continuous-Time Nonlinear Systems with Output Constraints

Approximate dynamic programming for continuous state and control problems

Optimal Control of Discrete-Time Nonlinear Systems