Off-Policy Reinforcement Learning For H-Infinity Control Design - Citegraph

Paper Info

Title
Off-Policy Reinforcement Learning For H-Infinity Control Design

Abstract
The H-infinity control design problem is considered for nonlinear systems with unknown internal system model. It is known that the nonlinear H-infinity control problem can be transformed into solving the so-called Hamilton-Jacobi-Isaacs (HJI) equation, which is a nonlinear partial differential equation that is generally impossible to be solved analytically. Even worse, model-based approaches cannot be used for approximately solving HJI equation, when the accurate system model is unavailable or costly to obtain in practice. To overcome these difficulties, an off-policy reinforcement leaning (RL) method is introduced to learn the solution of HJI equation from real system data instead of mathematical system model, and its convergence is proved. In the off-policy RL method, the system data can be generated with arbitrary policies rather than the evaluating policy, which is extremely important and promising for practical systems. For implementation purpose, a neural network (NN)-based actor-critic structure is employed and a least-square NN weight update algorithm is derived based on the method of weighted residuals. Finally, the developed NN-based off-policy RL method is tested on a linear F16 aircraft plant, and further applied to a rotational/translational actuator system.

Year	DOI	Venue
2013	10.1109/TCYB.2014.2319577	IEEE TRANSACTIONS ON CYBERNETICS
Keywords	DocType	Volume
H-infinity control design, Hamilton-Jacobi-Isaacs equation, neural network, off-policy learning, reinforcement learning	Journal	45
Issue	ISSN	Citations
1	2168-2267	84
PageRank	References	Authors
2.05	36	3

Authors (3 rows)

Cited by (84 rows)

References (36 rows)

Name	Order	Citations	PageRank
Biao Luo	1	554	23.80
Huai-Ning Wu	2	2104	98.52
Tingwen Huang	3	5684	310.24

1