Abstract
This article presents a novel lifelong integral reinforcement learning (LIRL)-based optimal trajectory tracking scheme using the multilayer (MNN) or deep neural network (Deep NN) for the uncertain nonlinear continuous-time (CT) affine systems subject to state constraints. A critic MNN, which approximates the value function, and a second NN identifier are together used to generate the optimal control policies. The weights of the critic MNN are tuned online using a novel singular value decomposition (SVD)-based method, which can be extended to MNN with the N-hidden layers. Moreover, an online lifelong learning (LL) scheme is incorporated with the critic MNN to mitigate the problem of catastrophic forgetting in the multitasking systems. Additionally, the proposed optimal framework addresses state constraints by utilizing a time-varying barrier function (TVBF). The uniform ultimate boundedness (UUB) of the overall closed-loop system is shown using the Lyapunov stability analysis. A two-link robotic manipulator that compares to recent literature shows a 47% total cost reduction, demonstrating the effectiveness of the proposed method.
Published Version
Talk to us
Join us for a 30 min session where you can share your feedback and ask us any queries you have