Real-Time Policy Optimization for UAV Swarms Based on Evolution Strategies

Zeyu Chen,Haiying Liu,Guohua Liu

doi:10.3390/drones8110619

Abstract

Multi-agent decision-making faces many challenges such as non-stationarity and sparse rewards, while the complexity and randomness of the real environment further complicate policy development. This paper addresses the high-dimensional policy optimization problems of unmanned aerial vehicle (UAV) swarms. By modeling the problem scenario as a Markov decision process, a real-time policy optimization algorithm based on evolution strategy (ES) pre-training is proposed. This approach combines decision-time planning with background planning to evaluate and integrate different sets of policy parameters in a temporal context. In the experimental phase, the policy network is trained using both ES and REINFORCE algorithms on a constructed simulation platform. Comparative experiments demonstrate the effectiveness of using ES for policy pre-training. Finally, the proposed real-time policy optimization algorithm further improves the performance of the swarm by approximately 10% in simulations, offering a feasible solution for adversarial games between swarms and extending the research scope of evolutionary algorithms.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Real-Time Policy Optimization for UAV Swarms Based on Evolution Strategies

Abstract

Talk to us

Similar Papers

More From: Drones

Lead the way for us

Journal: Drones	Publication Date: Oct 29, 2024
License type: CC BY 4.0

Similar Papers

Swarm of autonomous unmanned aerial vehicles with 3D deconfliction
Zbigniew Bogdanowicz
-
Zbigniew BogdanowiczZbigniew Bogdanowicz
09 May 2018
09 May 2018

Dual-game based UAV swarm obstacle avoidance algorithm in multi-narrow type obstacle scenarios
Ye Lin ... Yun Lin
EURASIP Journal on Advances in Signal Processing | VOL. 2023
Ye Lin, et. al.Ye Lin ... Yun Lin
16 Nov 2023
EURASIP Journal on Advances in Signal Processing | VOL. 2023

Collision Avoidance Method for Self-Organizing Unmanned Aerial Vehicle Flights
Yang Huang ... Jun Tang
IEEE Access | VOL. 7
Yang Huang, et. al.Yang Huang ... Jun Tang
01 Jan 2019
IEEE Access | VOL. 7

Adaptive decision-making with deep Q-network for heterogeneous unmanned aerial vehicle swarms in dynamic environments
Wenjia Su ... Dan Fang
Computers and Electrical Engineering | VOL. 119
Wenjia Su, et. al.Wenjia Su ... Dan Fang
07 Sep 2024
Computers and Electrical Engineering | VOL. 119

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Real-Time Policy Optimization for UAV Swarms Based on Evolution Strategies

Abstract

Talk to us

Similar Papers

More From: Drones