Reinforcement Learning With Task Decomposition for Cooperative Multiagent Systems.

Changyin Sun,Lu Dong,Wenzhang Liu

doi:10.1109/tnnls.2020.2996209

Abstract

In this article, we study cooperative multiagent systems (MASs) with multiple tasks by using reinforcement learning (RL)-based algorithms. The target for a single-agent RL system is represented by its scalar reward signals. However, for an MAS with multiple cooperative tasks, the holistic reward signal consists of multiple parts to represent the tasks, which makes the problem complicated. Existing multiagent RL algorithms search distributed policies with holistic reward signals directly, making it difficult to obtain an optimal policy for each task. This article provides efficient learning-based algorithms such that each agent can learn a joint optimal policy to accomplish these multiple tasks cooperatively with other agents. The main idea of the algorithms is to decompose the holistic reward signal for each agent into multiple parts according to the subtasks, and then the proposed algorithms learn multiple value functions with the decomposed reward signals and update the policy with the sum of distributed value functions. In addition, this article presents a theoretical analysis of the proposed approach. Finally, the simulation results for both discrete decision-making and continuous control problems have demonstrated the effectiveness of the proposed algorithms.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Reinforcement Learning With Task Decomposition for Cooperative Multiagent Systems.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems

Lead the way for us

Journal: IEEE Transactions on Neural Networks and Learning Systems	Publication Date: Jun 17, 2020
Citations: 79

Similar Papers

A step toward a unified treatment of continuous and discrete time control problems
Volker Mehrmann
Linear Algebra and its Applications | VOL. 241
Volker MehrmannVolker Mehrmann
01 Jul 1996
Linear Algebra and its Applications | VOL. 241

Improving Safety in Deep Reinforcement Learning using Unsupervised Action Planning
Hao-Lun Hsu ... Sehoon Ha
-
Hao-Lun Hsu, et. al.Hao-Lun Hsu ... Sehoon Ha
23 May 2022
23 May 2022

Decentralized Bayesian reinforcement learning for online agent collaboration
...
-
, et. al. ...
16 May 2018
16 May 2018

Computational Nonlinear Stochastic Control
Mrinal Kumar ... S Chakravorty
Journal of Guidance, Control, and Dynamics | VOL. 32
Mrinal Kumar, et. al.Mrinal Kumar ... S Chakravorty
01 May 2009
Journal of Guidance, Control, and Dynamics | VOL. 32

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Reinforcement Learning With Task Decomposition for Cooperative Multiagent Systems.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems