Multi-Task Deep Reinforcement Learning for Intelligent Logistics Path Planning and Scheduling Optimization

Xianfeng Zhu

doi:10.31449/inf.v49i20.7996

Contact Editors Europe, Africa:
Matjaz Gams
N. and S. America:
Karthick Gunasekaran
Asia, Australia:
Vinay Singh
Overview papers:
Maria Ganzha
Wiesław Pawlowski
Aleksander Denisiuk Abstacting / Indexing

Informatica is surveyed by:

ACM Digital Library
Citeseer
COBISS
Compendex
Computer & Information Systems Abstracts
Computer Database
Computer Science Index
dLib.si
DBLP Computer Science Bibliography
Directory of Open Access Journals
Google Scholar
InfoTrac OneFile
Inspec
Linguistic and Language Behaviour Abstracts
Mathematical Reviews, MatSciNet, MatSci on SilverPlatter and Current Mathematical Publications
Scopus Publishing

Informatica is published by:

Support

Informatica is supported by:

ACM Slovenia
Slovenian Society for Pattern Recognition
Slovenian Artificial Intelligence Society
Slovenian Society for Cognitive Science
Slovenian Society of Mathematicians, Physicists and Astronomers
Automatic Control Society of Slovenia
Slovenian Academy of Engineering
International Federation for Information Processing

Journal Help

User

Journal Content Search
Browse

Information

Notifications

About The Author

Xianfeng Zhu
School of Economics and Management, Jiaozuo University
China

Support & Indexing

Multi-Task Deep Reinforcement Learning for Intelligent Logistics Path Planning and Scheduling Optimization

Xianfeng Zhu

Abstract

Intelligent logistics systems optimize path planning and scheduling problems by introducing deep reinforcement learning (DRL) to cope with the complexity of dynamic demand and resource constraints. This study proposes an improved strategy that combines multi-task learning (MTL) with Q-learning to achieve simultaneous optimization of multiple related subtasks, such as path selection, task scheduling, and resource allocation, and shares some network parameters to improve model efficiency and generalization ability. The experimental design covers a variety of logistics scenarios, including urban distribution, long-distance transportation, and emergency response. Five baseline models are used for comparative evaluation to verify the advantages of the DRL method in computational efficiency, cost control, resource utilization, service level, and dynamic adaptability. Through experimental verification, the DRL model achieved a 15% cost reduction in cost control compared to traditional algorithms; the resource utilization rate reached 85%, and it performed excellently in terms of efficiency improvement. It has excellent adaptability when dealing with dynamic demands and can respond quickly to environmental changes, effectively improving the overall performance of the intelligent logistics system. It is particularly suitable for application scenarios that require real-time decision-making and support dynamic demand fluctuations.

Full Text:

PDF

DOI: https://doi.org/10.31449/inf.v49i20.7996

This work is licensed under a Creative Commons Attribution 3.0 License.

Informatica is financially supported by the Slovenian research agency from the Call for co-financing of scientific periodical publications.

Webmaster: Mario Konecki

Username
Password
Remember me