Frontiers in Emerging Multidisciplinary Sciences

Open Access Peer Review International
Open Access

Deep Reinforcement Learning Approach for Optimal Automatic Driving of Urban Rail Trains

4 Department of Artificial Intelligence Engineering Slovenian Center for Intelligent Systems Ljubljana, Slovenia
4 Center Laboratory of Machine Learning and Data Science Institute for Digital Innovation Maribor, Slovenia

Abstract

Automatic train operation requires a control strategy capable of maintaining speed-tracking accuracy, passenger comfort, operational safety, and energy efficiency under continuously changing railway conditions. Conventional control approaches can provide effective regulation, but their performance may be constrained when train dynamics, disturbances, operational constraints, and nonlinear relationships become difficult to represent through fixed control rules. Reinforcement learning provides an alternative paradigm in which an intelligent controller learns operational decisions through interaction with an environment. This study develops a conceptual deep reinforcement learning approach for optimal automatic driving of urban rail trains, with particular emphasis on Deep Q-Network (DQN)-based decision-making. The proposed approach integrates train-state representation, discrete driving-action selection, reward-based optimization, and operational constraints into a unified automatic driving framework. The literature indicates that iterative learning control, self-anti-disturbance control, Q-learning, policy-gradient reinforcement learning, and comfort-evaluation methods provide important foundations for intelligent train control. However, these approaches also reveal a need for a more integrated framework capable of balancing speed tracking, energy consumption, and ride comfort. The proposed methodology therefore formulates automatic driving as a sequential decision-making problem and defines a multi-objective reward mechanism incorporating tracking error, acceleration variation, energy consumption, and operational constraints. The resulting framework provides a theoretical basis for adaptive and optimization-oriented automatic train driving and establishes directions for future experimental validation using real or simulated urban rail operating data.

How to Cite

Dr. Luka Kranjc, & Dr. Maja Zupan. (2026). Deep Reinforcement Learning Approach for Optimal Automatic Driving of Urban Rail Trains. Frontiers in Emerging Multidisciplinary Sciences, 3(08), 56–61. Retrieved from https://irjernet.com/index.php/fems/article/view/497

References

S. H. Lai, C. J. Chen, L. Yan and Y. P. Li, “A comprehensive comfort evaluation system for subway train passengers based on hierarchical analysis,” Science Technology and Engineering, vol. 19, no. 36, pp. 296–301, 2019.
W. B. Lian, B. H. Liu, W. W. Li, X. Q. Liu, F. Y. Gao et al., “Automatic speed control of high-speed trains based on self-anti-disturbance control,” Journal of the China Railway Society, vol. 42, no. 1, pp. 76–81, 2020.
M. Zhang, “Research on automatic train driving method based on reinforcement learning,” Ph.D. dissertation, China Academy of Railway Science, China, 2020.
M. Zhang, Q. Zhang and Z. X. Zhang, “Research on energy-saving optimization of high-speed railroad trains based on Q-learning algorithm,” Railway Transportation and Economy, vol. 41, no. 12, pp. 111–117, 2019.
M. Zhang, Q. Zhang, W. T. Liu and B. Y. Zhou, “A train intelligent control method based on policy gradient reinforcement learning,” Journal of the China Railway Society, vol. 42, no. 1, pp. 69–75, 2020.
J. H. Wu and X. H. Zhang, “Comprehensive evaluation of ride comfort of urban rail trains based on fuzzy reasoning,” Journal of Zhejiang Normal University (Natural Sciences), vol. 40, no. 4, pp. 453–458, 2017.
Z. Y. He, “Application of adaptive iterative learning control in automatic train driving system,” Ph.D. dissertation, China Academy of Railway Science, China, 2019.
Z. Y. He and N. Xu, “Non-parametric iterative learning control algorithm for automatic train driving control,” Journal of the China Railway Society, vol. 42, no. 12, pp. 90–96, 2020.
J. Yang, Y. Q. Chen and P. P. Wang, “Self-anti-disturbance controller design for train speed tracking based on improved particle swarm algorithm,” Journal of the China Railway Society, vol. 43, no. 7, pp. 40–46, 2021.
K. Ramamurthy, R. K. Konduru and N. Amanmadov, "EvoGraphCoder: An Evolutionary Graph-Reasoning Framework for Self-Adaptive Software Engineering," in IEEE Access, vol. 14, pp. 63063-63076, 2026, doi: 10.1109/ACCESS.2026.3686019.
Geo Philip, Paulson, Integrated Intelligent Building Energy Management: A Multi-LayerFramework for Renewable Energy, HVAC Optimization, and SmartElectrical Network Coordination. Available at SSRN: https://ssrn.com/abstract=6993209 orhttp://dx.doi.org/10.2139/ssrn.6993209