arXiv · 2508.08493
POMO+: Leveraging starting nodes in POMO for solving Capacitated Vehicle Routing Problem
Abstract
In recent years, reinforcement learning (RL) methods have emerged as a promising approach for solving combinatorial problems. Among RL-based models, POMO has demonstrated strong performance on a variety of tasks, including variants of the Vehicle Routing Problem (VRP). However, there is room for improvement for these tasks. In this work, we improved POMO, creating a method (\textbf{POMO+}) that leverages the initial nodes to find a solution in a more informed way. We ran experiments on our new model and observed that our solution converges faster and achieves better results. We validated our models on the CVRPLIB dataset and noticed improvements in problem instances with up to 100 customers. We hope that our research in this project can lead to further advancements in the field.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Szymon Jakubicz, Karol Kuźniak, Jan Wawszczak, Paweł Gora. 2025-08-11. POMO+: Leveraging starting nodes in POMO for solving Capacitated Vehicle Routing Problem. https://arxiv.org/abs/2508.08493
Cite the original work for its findings. Save a collection to share your selection of sources.