A Decentralised Multi-Agent Reinforcement Learning Approach for the Same-Day Delivery Problem

03/22/2022
by   Elvin Ngu, et al.
0

Same-Day Delivery services are becoming increasingly popular in recent years. These have been usually modelled by previous studies as a certain class of Dynamic Vehicle Routing Problem (DVRP) where goods must be delivered from a depot to a set of customers in the same day that the orders were placed. Adaptive exact solution methods for DVRPs can become intractable even for small problem instances. In this paper, we formulate the SDDP as a Markov Decision Process (MDP) and solve it using a parameter-sharing Deep Q-Network, which corresponds to a decentralised Multi-Agent Reinforcement Learning (MARL) approach. For this, we create a multi-agent grid-based SDD environment, consisting of multiple vehicles, a central depot and dynamic order generation. In addition, we introduce zone-specific order generation and reward probabilities. We compare the performance of our proposed MARL approach against a Mixed Inter Programming (MIP) solution. Results show that our proposed MARL framework performs on par with MIP-based policy when the number of orders is relatively low. For problem instances with higher order arrival rates, computational results show that the MARL approach underperforms the MIP by up to 30 zone-specific parameters are employed. The gap is reduced from 30 5x5 grid scenario with 30 orders. Execution time results indicate that the MARL approach is, on average, 65 times faster than the MIP-based policy, and therefore may be more advantageous for real-time control, at least for small-sized instances.

READ FULL TEXT
research
05/12/2023

Multi-Agent Reinforcement Learning for Network Routing in Integrated Access Backhaul Networks

We investigate the problem of wireless routing in integrated access back...
research
04/24/2021

A Deep Reinforcement Learning Approach for the Meal Delivery Problem

We consider a meal delivery service fulfilling dynamic customer requests...
research
10/07/2019

Multi-Agent Reinforcement Learning for Order-dispatching via Order-Vehicle Distribution Matching

Improving the efficiency of dispatching orders to vehicles is a research...
research
04/17/2023

Control and Coordination of a SWARM of Unmanned Surface Vehicles using Deep Reinforcement Learning in ROS

An unmanned surface vehicle (USV) can perform complex missions by contin...
research
02/13/2020

Multi-Vehicle Routing Problems with Soft Time Windows: A Multi-Agent Reinforcement Learning Approach

Multi-vehicle routing problem with soft time windows (MVRPSTW) is an ind...
research
02/22/2022

A Framework for Multi-stage Bonus Allocation in meal delivery Platform

Online meal delivery is undergoing explosive growth, as this service is ...
research
01/05/2023

Playing hide and seek: tackling in-store picking operations while improving customer experience

The evolution of the retail business presents new challenges and raises ...

Please sign up or login with your details

Forgot password? Click here to reset