Intelligent Scheduling Method for Bulk Cargo Terminal Loading Process Based on Deep Reinforcement Learning

被引:3
|
作者
Li, Changan [1 ,2 ]
Wu, Sirui [3 ]
Li, Zhan [3 ,4 ]
Zhang, Yuxiao [3 ]
Zhang, Lijie [1 ]
Gomes, Luis [5 ]
机构
[1] Yanshan Univ, Key Lab Adv Forging & Stamping Technol & Sci, Minist Educ China, Qinhuangdao 066004, Hebei, Peoples R China
[2] Chnenergy Tianjin Port Co Ltd, Tianjin 300450, Peoples R China
[3] Harbin Inst Technol, Res Inst Intelligent Control & Syst, Harbin 150001, Peoples R China
[4] Ningbo Inst Intelligent Equipment Technol Co Ltd, Ningbo 315201, Peoples R China
[5] NOVA Univ Lisbon, NOVA Sch Sci & Technol, Ctr Technol & Syst, P-2829516 Monte De Caparica, Portugal
基金
中国国家自然科学基金;
关键词
bulk cargo loading; MDP model; deep reinforcement learning; intelligent scheduling; BERTH ALLOCATION PROBLEM; OPTIMIZATION; SYSTEM;
D O I
10.3390/electronics11091390
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Sea freight is one of the most important ways for the transportation and distribution of coal and other bulk cargo. This paper proposes a method for optimizing the scheduling efficiency of the bulk cargo loading process based on deep reinforcement learning. The process includes a large number of states and possible choices that need to be taken into account, which are currently performed by skillful scheduling engineers on site. In terms of modeling, we extracted important information based on actual working data of the terminal to form the state space of the model. The yard information and the demand information of the ship are also considered. The scheduling output of each convey path from the yard to the cabin is the action of the agent. To avoid conflicts of occupying one machine at same time, certain restrictions are placed on whether the action can be executed. Based on Double DQN, an improved deep reinforcement learning method is proposed with a fully connected network structure and selected action sets according to the value of the network and the occupancy status of environment. To make the network converge more quickly, an improved new epsilon-greedy exploration strategy is also proposed, which uses different exploration rates for completely random selection and feasible random selection of actions. After training, an improved scheduling result is obtained when the tasks arrive randomly and the yard state is random. An important contribution of this paper is to integrate the useful features of the working time of the bulk cargo terminal into a state set, divide the scheduling process into discrete actions, and then reduce the scheduling problem into simple inputs and outputs. Another major contribution of this article is the design of a reinforcement learning algorithm for the bulk cargo terminal scheduling problem, and the training efficiency of the proposed algorithm is improved, which provides a practical example for solving bulk cargo terminal scheduling problems using reinforcement learning.
引用
收藏
页数:18
相关论文
共 50 条
  • [1] An improved deep reinforcement learning approach: A case study for optimisation of berth and yard scheduling for bulk cargo terminal
    Ai, T.
    Huang, L.
    Song, R. J.
    Huang, H. F.
    Jiao, F.
    Ma, W. G.
    [J]. ADVANCES IN PRODUCTION ENGINEERING & MANAGEMENT, 2023, 18 (03): : 303 - 316
  • [2] Yard Crane Scheduling Method Based on Deep Reinforcement Learning for the Automated Container Terminal
    Wang, Wuyin
    Huang, Zizhao
    Zhuang, Zilong
    Fang, Huaijin
    Qin, Wei
    [J]. Jixie Gongcheng Xuebao/Journal of Mechanical Engineering, 2024, 60 (06): : 44 - 57
  • [3] Intelligent bulk cargo terminal scheduling based on a novel chaotic-optimal thermodynamic evolutionary algorithm
    Liu, Shida
    Liu, Qingsheng
    Wang, Li
    Chen, Xianlong
    [J]. COMPLEX & INTELLIGENT SYSTEMS, 2024, 10 (06) : 7435 - 7450
  • [4] An intelligent monitoring method of underground unmanned electric locomotive loading process based on deep learning method
    Ren, Zhu-li
    Zhang, Jin-long
    Yuan, Rui-fu
    [J]. COGENT ENGINEERING, 2024, 11 (01):
  • [5] Intelligent deep reinforcement learning-based scheduling in relay-based HetNets
    Chen, Chao
    Wu, Zhengyang
    Yu, Xiaohan
    Ma, Bo
    Li, Chuanhuang
    [J]. EURASIP JOURNAL ON WIRELESS COMMUNICATIONS AND NETWORKING, 2023, 2023 (01)
  • [6] Intelligent deep reinforcement learning-based scheduling in relay-based HetNets
    Chao Chen
    Zhengyang Wu
    Xiaohan Yu
    Bo Ma
    Chuanhuang Li
    [J]. EURASIP Journal on Wireless Communications and Networking, 2023
  • [7] Resources Scheduling for Ambient Backscatter Communication-Based Intelligent IIoT: A Collective Deep Reinforcement Learning Method
    Huang, Yudian
    Li, Meng
    Yu, F. Richard
    Si, Pengbo
    Zhang, Haijun
    Qiao, Junfei
    [J]. IEEE TRANSACTIONS ON COGNITIVE COMMUNICATIONS AND NETWORKING, 2024, 10 (02) : 634 - 648
  • [8] Research on Robot Intelligent Control Method Based on Deep Reinforcement Learning
    Rao, Shu
    [J]. 2022 6TH INTERNATIONAL SYMPOSIUM ON COMPUTER SCIENCE AND INTELLIGENT CONTROL, ISCSIC, 2022, : 221 - 225
  • [9] Adaptive Clutter Intelligent Suppression Method Based on Deep Reinforcement Learning
    Cheng, Yi
    Su, Junjie
    Xiu, Chunbo
    Liu, Jiaxin
    [J]. APPLIED SCIENCES-BASEL, 2024, 14 (17):
  • [10] State Evaluation Method of Distribution Terminal Based on Deep Reinforcement Learning
    Xue, Fei
    Li, Xutao
    Wang, Xiaoli
    Li, Hongqiang
    Tian, Bei
    [J]. MATHEMATICAL PROBLEMS IN ENGINEERING, 2022, 2022