Multi-Decoder Attention Model with Embedding Glimpse for Solving Vehicle Routing Problems

About

We present a novel deep reinforcement learning method to learn construction heuristics for vehicle routing problems. In specific, we propose a Multi-Decoder Attention Model (MDAM) to train multiple diverse policies, which effectively increases the chance of finding good solutions compared with existing methods that train only one policy. A customized beam search strategy is designed to fully exploit the diversity of MDAM. In addition, we propose an Embedding Glimpse layer in MDAM based on the recursive nature of construction, which can improve the quality of each policy by providing more informative embeddings. Extensive experiments on six different routing problems show that our method significantly outperforms the state-of-the-art deep learning based models.

Liang Xin, Wen Song, Zhiguang Cao, Jie Zhang• 2020

Related benchmarks

Task	Dataset	Result
Traveling Salesman Problem (TSP)	TSP n=100 10K instances (test)	Objective Value7.79	52
Capacitated Vehicle Routing Problem	CVRP N=100 10,000 instances (test)	Objective Value15.99	44
Traveling Salesperson Problem	TSP-100	Solution Length7.79	42
Traveling Salesperson Problem	TSP N=500 Generalization (128 instances)	Optimality Gap9.88	41
Capacitated Vehicle Routing Problem	CVRP N=20 10,000 instances (test)	Objective Value6.14	38
Traveling Salesperson Problem	TSP N=200 (Generalization (128 instances))	Optimality Gap2.04	35
Traveling Salesperson Problem	TSP N=100 (test)	Optimality Gap0.39	33
Traveling Salesperson Problem	TSP N=1000 Generalization (128 instances)	Optimality Gap19.96	30
Capacitated Vehicle Routing Problem	CVRP N=50 10,000 instances (test)	Objective Value10.48	29
Traveling Salesman Problem	Euclidean TSP N=50	Optimal Tour Length5.7	26

Showing 10 of 24 rows

Other info

Follow for update

@wizwand_team Discord