Decentralised Multi-Agent Reinforcement Learning Approach for the Same-Day Delivery Problem

Elvin Ngu,Jose Javier Escribano Macias,Leandro Parada,Panagiotis Angeloudis

doi:10.1177/03611981221093324

Abstract

Same-day delivery (SDD) services have become increasingly popular in recent years. These have been usually modeled by previous studies as a certain class of dynamic vehicle routing problem (DVRP) where goods must be delivered from a depot to a set of customers in the same day that the orders were placed. Adaptive exact solution methods for DVRPs can become intractable even for small problem instances. In this paper, the same-day delivery problem (SDDP) is formulated as a Markov decision process (MDP) and it is solved using a parameter-sharing Deep Q-Network, which corresponds to a decentralised multi-agent reinforcement learning (MARL) approach. For this, a multi-agent grid-based SDD environment is created, consisting of multiple vehicles, a central depot, and dynamic order generation. In addition, zone-specific order generation and reward probabilities are introduced. The performance of the proposed MARL approach is compared against a mixed-integer programming (MIP) solution. Results show that the proposed MARL framework performs on par with MIP-based policy when the number of orders is relatively low. For problem instances with higher order arrival rates, computational results show that the MARL approach underperforms MIP by up to 30%. The performance gap between both methods becomes smaller when zone-specific parameters are employed. The gap is reduced from 30% to 3% for a 5 × 5 grid scenario with 30 orders. Execution time results indicate that the MARL approach is, on average, 65 times faster than the MIP-based policy, and therefore may be more advantageous for real-time control, at least for small-sized instances.

Full Text

Published version (

Free)

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Transportation Research Record: Journal of the Transportation Research Board	Publication Date: Jun 23, 2022
Citations: 1	License type: CC BY-NC 4.0

R Discovery Prime

R Discovery Prime

Decentralised Multi-Agent Reinforcement Learning Approach for the Same-Day Delivery Problem

Abstract

Talk to us

Similar Papers

More From: Transportation Research Record: Journal of the Transportation Research Board

Lead the way for us

Similar Papers

Engineering A Large-Scale Traffic Signal Control: A Multi-Agent Reinforcement Learning Approach
Guoqiang Mao ... Changle Li
-
Guoqiang Mao, et. al.Guoqiang Mao ... Changle Li
10 May 2021
10 May 2021

A Multi-Agent Reinforcement Learning Approach to Traffic Control at Future Urban Air Mobility Intersections
Sabrullah Deniz ... Zhenbo Wang
-
Sabrullah Deniz, et. al.Sabrullah Deniz ... Zhenbo Wang
03 Jan 2022
03 Jan 2022

On the Role of Reward Functions for Reinforcement Learning in the Traffic Assignment Problem
Gabriel De Oliveira Ramos ... Ricardo Grunitzki
-
Gabriel De Oliveira Ramos, et. al.Gabriel De Oliveira Ramos ... Ricardo Grunitzki
01 Jul 2020
01 Jul 2020

Flexible Formation Control Using Hausdorff Distance: A Multi-agent Reinforcement Learning Approach
Yuzi Yan ... Zexu Zhang
-
Yuzi Yan, et. al.Yuzi Yan ... Zexu Zhang
29 Aug 2022
29 Aug 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Decentralised Multi-Agent Reinforcement Learning Approach for the Same-Day Delivery Problem

Abstract

Talk to us

Similar Papers

More From: Transportation Research Record: Journal of the Transportation Research Board