Reinforced Lin-Kernighan-Helsgaun Algorithms for the Traveling Salesman Problems
Abstract
TSP is a classical NP-hard combinatorial optimization problem with many practical variants. LKH is one of the state-of-the-art local search algorithms for the TSP. LKH-3 is a powerful extension of LKH that can solve many TSP variants. Both LKH and LKH-3 associate a candidate set to each city to improve the efficiency, and have two different methods, -measure and POPMUSIC, to decide the candidate sets. In this work, we first propose a Variable Strategy Reinforced LKH (VSR-LKH) algorithm, which incorporates three reinforcement learning methods (Q-learning, Sarsa, Monte Carlo) with LKH, for the TSP. We further propose a new algorithm called VSR-LKH-3 that combines the variable strategy reinforcement learning method with LKH-3 for typical TSP variants, including the TSP with time windows (TSPTW) and Colored TSP (CTSP). The proposed algorithms replace the inflexible traversal operations in LKH and LKH-3 and let the algorithms learn to make a choice at each search step by reinforcement learning. Both LKH and LKH-3, with either -measure or POPMUSIC, can be significantly improved by our methods. Extensive experiments on 236 widely-used TSP benchmarks with up to 85,900 cities demonstrate the excellent performance of VSR-LKH. VSR-LKH-3 also significantly outperforms the state-of-the-art heuristics for TSPTW and CTSP.
Cite
@article{arxiv.2207.03876,
title = {Reinforced Lin-Kernighan-Helsgaun Algorithms for the Traveling Salesman Problems},
author = {Jiongzhi Zheng and Kun He and Jianrong Zhou and Yan Jin and Chu-Min Li},
journal= {arXiv preprint arXiv:2207.03876},
year = {2022}
}
Comments
arXiv admin note: text overlap with arXiv:2107.06870