Optimal tracking control of nonlinear partially-unknown constrained-input systems using integral reinforcement learning

被引:516
作者
Modares, Hamidreza [1 ]
Lewis, Frank L. [1 ]
机构
[1] Univ Texas Arlington, Res Inst, Ft Worth, TX 76118 USA
关键词
Optimal tracking control; Integral reinforcement learning; Input constrainers; Neural networks; ADAPTIVE OPTIMAL-CONTROL; POLICY ITERATION; TIME-SYSTEMS; APPROXIMATION;
D O I
10.1016/j.automatica.2014.05.011
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
In this paper, a new formulation for the optimal tracking control problem (OTCP) of continuous-time nonlinear systems is presented. This formulation extends the integral reinforcement learning (IRL) technique, a method for solving optimal regulation problems, to learn the solution to the OTCP. Unlike existing solutions to the OTCP, the proposed method does not need to have or to identify knowledge of the system drift dynamics, and it also takes into account the input constraints a priori. An augmented system composed of the error system dynamics and the command generator dynamics is used to introduce a new nonquadratic discounted performance function for the OTCP. This encodes the input constrains into the optimization problem. A tracking Hamilton-Jacobi-Bellman (HJB) equation associated with this nonquadratic performance function is derived which gives the optimal control solution. An online IRL algorithm is presented to learn the solution to the tracking HJB equation without knowing the system drift dynamics. Convergence to a near-optimal control solution and stability of the whole system are shown under a persistence of excitation condition. Simulation examples are provided to show the effectiveness of the proposed method. (C) 2014 Elsevier Ltd. All rights reserved.
引用
收藏
页码:1780 / 1792
页数:13
相关论文
共 33 条
  • [21] Sastry S., 1991, Nonlinear systems: analysis, stability, and control
  • [22] Slotine J.-J. E., 1991, Applied nonlinear control
  • [23] Sutton RS, 2018, ADAPT COMPUT MACH LE, P1
  • [24] Vamvoudakis K., 2013, INT J ROBUST NONLINE
  • [25] Online actor-critic algorithm to solve the continuous-time infinite horizon optimal control problem
    Vamvoudakis, Kyriakos G.
    Lewis, Frank L.
    [J]. AUTOMATICA, 2010, 46 (05) : 878 - 888
  • [26] Adaptive optimal control for continuous-time linear systems based on policy iteration
    Vrabie, D.
    Pastravanu, O.
    Abu-Khalaf, M.
    Lewis, F. L.
    [J]. AUTOMATICA, 2009, 45 (02) : 477 - 484
  • [27] Neural network approach to continuous-time direct adaptive optimal control for partially unknown nonlinear systems
    Vrabie, Draguna
    Lewis, Frank
    [J]. NEURAL NETWORKS, 2009, 22 (03) : 237 - 246
  • [28] Finite-horizon neuro-optimal tracking control for a class of discrete-time nonlinear systems using adaptive dynamic programming approach
    Wang, Ding
    Liu, Derong
    Wei, Qinglai
    [J]. NEUROCOMPUTING, 2012, 78 (01) : 14 - 22
  • [29] Werbos P., 1992, HDB INTELLIGENT CONT, P493
  • [30] NEURAL NETWORKS FOR CONTROL AND SYSTEM-IDENTIFICATION
    WERBOS, PJ
    [J]. PROCEEDINGS OF THE 28TH IEEE CONFERENCE ON DECISION AND CONTROL, VOLS 1-3, 1989, : 260 - 265