Inverse Reinforcement Learning Control for Linear Multiplayer Games

Bosen Lian,Tianyou Chai,Vrushabh S Donge,Frank L Lewis,Ali Davoudi

doi:10.1109/cdc51059.2022.9993367

Abstract

This paper proposes model-based and model-free inverse reinforcement learning (RL) control algorithms for multiplayer game systems described by linear continuous-time differential equations. Both algorithms find the learner the same optimal control policies and trajectories as the expert, by inferring the unknown expert players’ cost functions from the expert’s trajectories. This paper first discusses a model-based inverse RL policy iteration that consists of 1) policy evaluation for cost matrices using a Lyapunov equation, 2) state-reward weight improvement using inverse optimal control (IOC), and 3) policy improvement using optimal control. Based on the model-based algorithm, an online data-driven inverse RL algorithm is proposed without knowing system dynamics or expert control gains. Rigorous convergence and stability analysis of these algorithms are provided. Finally, a simulation example verifies our approach.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Inverse Reinforcement Learning Control for Linear Multiplayer Games

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Data-Driven Inverse Reinforcement Learning Control for Linear Multiplayer Games.
Bosen Lian ... Vrushabh S Donge
IEEE Transactions on Neural Networks and Learning Systems | VOL. 35
Bosen Lian, et. al.Bosen Lian ... Vrushabh S Donge
01 Feb 2024
IEEE Transactions on Neural Networks and Learning Systems | VOL. 35

Inverse Reinforcement Learning for Adversarial Apprentice Games.
Bosen Lian ... Tianyou Chai
IEEE Transactions on Neural Networks and Learning Systems | VOL. 34
Bosen Lian, et. al.Bosen Lian ... Tianyou Chai
01 Aug 2023
IEEE Transactions on Neural Networks and Learning Systems | VOL. 34

Inverse reinforcement learning for multi-player noncooperative apprentice games
Bosen Lian ... Tianyou Chai
Automatica | VOL. 145
Bosen Lian, et. al.Bosen Lian ... Tianyou Chai
11 Aug 2022
Automatica | VOL. 145

Inferring Human-Robot Performance Objectives During Locomotion Using Inverse Reinforcement Learning and Inverse Optimal Control
Wentao Liu ... Jennie Si
IEEE Robotics and Automation Letters | VOL. 7
Wentao Liu, et. al.Wentao Liu ... Jennie Si
01 Apr 2022
IEEE Robotics and Automation Letters | VOL. 7

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Inverse Reinforcement Learning Control for Linear Multiplayer Games

Abstract

Talk to us

Similar Papers