Optimal tracking control of nonlinear partially-unknown constrained-input systems using integral reinforcement learning

被引：546

作者：

Modares, Hamidreza ^{[1
]}

Lewis, Frank L. ^{[1
]}

机构：

[1] Univ Texas Arlington, Res Inst, Ft Worth, TX 76118 USA

来源：

AUTOMATICA | 2014年 / 50卷 / 07期

关键词：

Optimal tracking control; Integral reinforcement learning; Input constrainers; Neural networks; ADAPTIVE OPTIMAL-CONTROL; POLICY ITERATION; TIME-SYSTEMS; APPROXIMATION;

D O I：

10.1016/j.automatica.2014.05.011

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

In this paper, a new formulation for the optimal tracking control problem (OTCP) of continuous-time nonlinear systems is presented. This formulation extends the integral reinforcement learning (IRL) technique, a method for solving optimal regulation problems, to learn the solution to the OTCP. Unlike existing solutions to the OTCP, the proposed method does not need to have or to identify knowledge of the system drift dynamics, and it also takes into account the input constraints a priori. An augmented system composed of the error system dynamics and the command generator dynamics is used to introduce a new nonquadratic discounted performance function for the OTCP. This encodes the input constrains into the optimization problem. A tracking Hamilton-Jacobi-Bellman (HJB) equation associated with this nonquadratic performance function is derived which gives the optimal control solution. An online IRL algorithm is presented to learn the solution to the tracking HJB equation without knowing the system drift dynamics. Convergence to a near-optimal control solution and stability of the whole system are shown under a persistence of excitation condition. Simulation examples are provided to show the effectiveness of the proposed method. (C) 2014 Elsevier Ltd. All rights reserved.

引用

页码：1780 / 1792

页数：13

共 33 条

[21]

Sastry S., 1991, Nonlinear systems: analysis, stability, and control

[22]

Slotine J.-J. E., 1991, Applied nonlinear control

[23]

Sutton RS, 2018, ADAPT COMPUT MACH LE, P1

[24]

Vamvoudakis K., 2013, INT J ROBUST NONLINE

[25] Online actor-critic algorithm to solve the continuous-time infinite horizon optimal control problem [J].

Vamvoudakis, Kyriakos G. ;

Lewis, Frank L. .

AUTOMATICA, 2010, 46 (05) :878-888

[26] Adaptive optimal control for continuous-time linear systems based on policy iteration [J].

Vrabie, D. ;

Pastravanu, O. ;

Abu-Khalaf, M. ;

Lewis, F. L. .

AUTOMATICA, 2009, 45 (02) :477-484

[27] Neural network approach to continuous-time direct adaptive optimal control for partially unknown nonlinear systems [J].

Vrabie, Draguna ;

Lewis, Frank .

NEURAL NETWORKS, 2009, 22 (03) :237-246

[28] Finite-horizon neuro-optimal tracking control for a class of discrete-time nonlinear systems using adaptive dynamic programming approach [J].

Wang, Ding ;

Liu, Derong ;

Wei, Qinglai .

NEUROCOMPUTING, 2012, 78 (01) :14-22

[29]

Werbos P., 1992, HDB INTELLIGENT CONT, P493

[30] NEURAL NETWORKS FOR CONTROL AND SYSTEM-IDENTIFICATION [J].

WERBOS, PJ .

PROCEEDINGS OF THE 28TH IEEE CONFERENCE ON DECISION AND CONTROL, VOLS 1-3, 1989, :260-265

← 1 2 3 4 →