Multi-Decoder Attention Model with Embedding Glimpse for Solving Vehicle Routing Problems

12/19/2020
by   Liang Xin, et al.
11

We present a novel deep reinforcement learning method to learn construction heuristics for vehicle routing problems. In specific, we propose a Multi-Decoder Attention Model (MDAM) to train multiple diverse policies, which effectively increases the chance of finding good solutions compared with existing methods that train only one policy. A customized beam search strategy is designed to fully exploit the diversity of MDAM. In addition, we propose an Embedding Glimpse layer in MDAM based on the recursive nature of construction, which can improve the quality of each policy by providing more informative embeddings. Extensive experiments on six different routing problems show that our method significantly outperforms the state-of-the-art deep learning based models.

READ FULL TEXT

page 7

page 11

research
12/12/2019

Learning Improvement Heuristics for Solving the Travelling Salesman Problem

Recent studies in using deep learning to solve the Travelling Salesman P...
research
03/04/2023

Neural Airport Ground Handling

Airport ground handling (AGH) offers necessary operations to flights dur...
research
10/06/2021

Learning to Iteratively Solve Routing Problems with Dual-Aspect Collaborative Transformer

Recently, Transformer has become a prevailing deep architecture for solv...
research
10/15/2021

NeuroLKH: Combining Deep Learning Model with Lin-Kernighan-Helsgaun Heuristic for Solving the Traveling Salesman Problem

We present NeuroLKH, a novel algorithm that combines deep learning with ...
research
12/24/2020

Learning Vehicle Routing Problems using Policy Optimisation

Deep reinforcement learning (DRL) has been used to learn effective heuri...
research
04/03/2020

Learning 2-opt Heuristics for the Traveling Salesman Problem via Deep Reinforcement Learning

Recent works using deep learning to solve the Traveling Salesman Problem...
research
05/31/2022

Sample-Efficient, Exploration-Based Policy Optimisation for Routing Problems

Model-free deep-reinforcement-based learning algorithms have been applie...

Please sign up or login with your details

Forgot password? Click here to reset