Consensus of Multi-agent Reinforcement Learning Systems: The Effect of Immediate Rewards

Neshat Elhami Fard,Rastko Selmic

doi:10.18196/jrc.v3i2.13082

Abstract

This paper studies the consensus problem of a leaderless, homogeneous, multi-agent reinforcement learning (MARL) system using actor-critic algorithms with and without malicious agents. The goal of each agent is to reach the consensus position with the maximum cumulative reward. Although the reward function converges in both scenarios, in the absence of the malicious agent, the cumulative reward is higher than with the malicious agent present. We consider here various immediate reward functions. First, we study the immediate reward function based on Manhattan distance. In addition to proposing three different immediate reward functions based on Euclidean, $n$-norm, and Chebyshev distances, we have rigorously shown which method has a better performance based on a cumulative reward for each agent and the entire team of agents. Finally, we present a combination of various immediate reward functions that yields a higher cumulative reward for each agent and the team of agents. By increasing the agents’ cumulative reward using the combined immediate reward function, we have demonstrated that the cumulative team reward in the presence of a malicious agent is comparable with the cumulative team reward in the absence of the malicious agent. The claims have been proven theoretically, and the simulation confirms theoretical findings.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Journal of Robotics and Control (JRC)	Publication Date: Feb 5, 2022
Citations: 5	License type: CC BY-SA 4.0

R Discovery Prime

R Discovery Prime

Consensus of Multi-agent Reinforcement Learning Systems: The Effect of Immediate Rewards

Abstract

Talk to us

Similar Papers

More From: Journal of Robotics and Control (JRC)

Lead the way for us

Similar Papers

Towards Efficient Coordination in Multi-Agent Reinforcement Learning through Hybrid Information-Driven Approaches
Priyanka S Chauhan
International Journal for Research in Applied Science and Engineering Technology | VOL. 12
Priyanka S ChauhanPriyanka S Chauhan
31 Mar 2024
International Journal for Research in Applied Science and Engineering Technology | VOL. 12

Formal Reachability Analysis for Multi-Agent Reinforcement Learning Systems
Xiaoyan Wang ... Shuqiu Li
IEEE Access | VOL. 9
Xiaoyan Wang, et. al.Xiaoyan Wang ... Shuqiu Li
01 Jan 2020
IEEE Access | VOL. 9

On the rationality of profit sharing in multi-agent reinforcement learning
K Miyazaki ... S Kobayashi
-
K Miyazaki, et. al.K Miyazaki ... S Kobayashi
30 Oct 2001
30 Oct 2001

Automated clash resolution for reinforcement steel design in concrete frames via Q-learning and Building Information Modeling
Jiepeng Liu ... Y Frank Chen
Automation in Construction | VOL. 112
Jiepeng Liu, et. al.Jiepeng Liu ... Y Frank Chen
31 Jan 2020
Automation in Construction | VOL. 112

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Consensus of Multi-agent Reinforcement Learning Systems: The Effect of Immediate Rewards

Abstract

Talk to us

Similar Papers

More From: Journal of Robotics and Control (JRC)