arXiv Machine Learning By Yida Xu, Zhaofang Mao, Yuheng Miao, Jiaxin Zhang, Yiting Sun

Integrated Order Dispatching and Routing for Last-Mile Pickup via Deep Reinforcement Learning

Read the original on arXiv Machine Learning →

arXiv:2607. 22356v1 Announce Type: new Abstract: In recent years, the growing complexity of last-mile pickup operations has increased the need for fast and accurate decision-making on logistics platforms.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Jul 21

A Deep Reinforcement Learning Algorithm for the Vehicle Routing Problem with Stochastic Demands and Outsourcing

arXiv:2607. 16875v1 Announce Type: cross Abstract: We introduce the vehicle routing problem with stochastic demands and outsourcing options (VRP-SDO), in which a logistics service provider partitions customer requests into customers outsourced to a common carrier and customers committed to its fixed fleet.

By Mohsen Dastpak, Fausto Errico, Ola Jabali
arXiv Machine Learning
Sep 17

Integrated Optimization of Automated Warehouse Operations and Last-Mile Transport for Differentiated On-Demand Delivery

The paper introduces an integrated optimization framework that links automated warehouse operations with last‑mile multi‑modal transport for differentiated on‑demand delivery. It employs a deep reinforcement learning approach—MORM‑AGDQN for warehouse scheduling and MRMH‑HCVRP for external routing—to balance service level, cost, and demand. The results demonstrate significant performance gains, including a 100 % on‑time delivery rate, a 29.3 % reduction in average last‑mile delivery time, a 46.4 % cut in total transportation distance, and a high‑priority service rate exceeding 92 % while maintaining cost‑customer satisfaction balance.

By Xiaozhu Sun, Bilal Farooq
arXiv AI
Aug 25

Memory-Enhanced Neural Solvers for Routing Problems

The paper introduces MEMENTO, a memory‑enhanced neural solver that improves routing problem solutions by using online data from repeated attempts to adjust action distributions during inference. It targets NP‑hard routing tasks such as the Traveling Salesman and Capacitated Vehicle Routing problems, outperforming existing tree‑search and policy‑gradient fine‑tuning methods. MEMENTO demonstrates strong scalability and data efficiency, achieving state‑of‑the‑art results on 11 of 12 evaluated tasks and enabling zero‑shot integration with diversity‑based solvers.

By Felix Chalumeau, Refiloe Shabe, Noah De Nicola, Arnu Pretorius, Thomas D. Barrett, Nathan Grinsztajn