IDEAS home Printed from https://ideas.repec.org/a/eee/transe/v200y2025ics1366554525001838.html

Large language model-enhanced reinforcement learning for generic bus holding control strategies

Author

Listed:
  • Yu, Jiajie
  • Wang, Yuhong
  • Ma, Wei

Abstract

Bus holding control is a widely-adopted strategy for maintaining stability and improving the operational efficiency of bus systems. Traditional model-based methods often face challenges with the low accuracy of bus state prediction and passenger demand estimation. In contrast, Reinforcement Learning (RL), as a data-driven approach, has demonstrated great potential in formulating bus holding strategies. RL determines the optimal control strategies in order to maximize the cumulative reward, which reflects the overall control goals. However, translating sparse and delayed control goals in real-world tasks into dense and real-time rewards for RL is challenging, normally requiring extensive manual trial-and-error. In view of this, this study introduces an automatic reward generation paradigm by leveraging the in-context learning and reasoning capabilities of Large Language Models (LLMs). This new paradigm, termed the LLM-enhanced RL, comprises several LLM-based modules: reward initializer, reward modifier, performance analyzer, and reward refiner. These modules cooperate to initialize and iteratively improve the reward function according to the feedback from training and test results for the specified RL-based task. Ineffective reward functions generated by the LLM are filtered out to ensure the stable evolution of the RL agents’ performance over iterations. To evaluate the feasibility of the proposed LLM-enhanced RL paradigm, it is applied to extensive bus holding control scenarios that vary in the number of bus lines, stops, and passenger demand. The results demonstrate the superiority, generalization capability, and robustness of the proposed paradigm compared to vanilla RL strategies, the LLM-based controller, physics-based feedback controllers, and optimization-based controllers. This study sheds light on the great potential of utilizing LLMs in various smart mobility applications.

Suggested Citation

  • Yu, Jiajie & Wang, Yuhong & Ma, Wei, 2025. "Large language model-enhanced reinforcement learning for generic bus holding control strategies," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 200(C).
  • Handle: RePEc:eee:transe:v:200:y:2025:i:c:s1366554525001838
    DOI: 10.1016/j.tre.2025.104142
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S1366554525001838
    Download Restriction: Full text for ScienceDirect subscribers only

    File URL: https://libkey.io/10.1016/j.tre.2025.104142?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to

    for a different version of it.

    References listed on IDEAS

    as
    1. Wang, Zhichao & Jiang, Rui & Jiang, Yu & Gao, Ziyou & Liu, Ronghui, 2024. "Modelling bus bunching along a common line corridor considering passenger arrival time and transfer choice under stochastic travel time," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 181(C).
    2. Weihua Gu & Michael J. Cassidy & Yuwei Li, 2015. "Models of Bus Queueing at Curbside Stops," Transportation Science, INFORMS, vol. 49(2), pages 204-212, May.
    3. Wu, Weitiao & Liu, Ronghui & Jin, Wenzhou, 2017. "Modelling bus bunching and holding control with vehicle overtaking and distributed passenger boarding behaviour," Transportation Research Part B: Methodological, Elsevier, vol. 104(C), pages 175-197.
    4. Dai, Zhuang & Liu, Xiaoyue Cathy & Chen, Zhuo & Guo, Renyong & Ma, Xiaolei, 2019. "A predictive headway-based bus-holding strategy with dynamic control point selection: A cooperative game theory approach," Transportation Research Part B: Methodological, Elsevier, vol. 125(C), pages 29-51.
    5. Lacombe, Rémi & Murgovski, Nikolce & Gros, Sébastien & Kulcsár, Balázs, 2024. "Integrated charging scheduling and operational control for an electric bus network," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 186(C).
    6. Lee, Enoch & Cen, Xuekai & Lo, Hong K., 2022. "Scheduling zonal-based flexible bus service under dynamic stochastic demand and Time-dependent travel time," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 168(C).
    7. Volodymyr Mnih & Koray Kavukcuoglu & David Silver & Andrei A. Rusu & Joel Veness & Marc G. Bellemare & Alex Graves & Martin Riedmiller & Andreas K. Fidjeland & Georg Ostrovski & Stig Petersen & Charle, 2015. "Human-level control through deep reinforcement learning," Nature, Nature, vol. 518(7540), pages 529-533, February.
    8. Xuan, Yiguang & Argote, Juan & Daganzo, Carlos F., 2011. "Dynamic bus holding strategies for schedule reliability: Optimal linear control and performance analysis," Transportation Research Part B: Methodological, Elsevier, vol. 45(10), pages 1831-1845.
    9. Hernández, Daniel & Muñoz, Juan Carlos & Giesen, Ricardo & Delgado, Felipe, 2015. "Analysis of real-time control strategies in a corridor with multiple bus services," Transportation Research Part B: Methodological, Elsevier, vol. 78(C), pages 83-105.
    10. Gkiotsalitis, K. & Cats, O., 2021. "At-stop control measures in public transport: Literature review and research agenda," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 145(C).
    11. Davide Castelvecchi, 2016. "Can we open the black box of AI?," Nature, Nature, vol. 538(7623), pages 20-23, October.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Minyu Shen & Weihua Gu & Michael J. Cassidy & Yongjie Lin & Wei Ni, 2024. "A vicious cycle along busy bus corridors and how to abate it," Papers 2403.08230, arXiv.org.
    2. Wang, Zhichao & Jiang, Rui, 2025. "Optimal bus holding control on bus routes traversing a common corridor," Physica A: Statistical Mechanics and its Applications, Elsevier, vol. 667(C).
    3. Li, Shukai & Liu, Ronghui & Yang, Lixing & Gao, Ziyou, 2019. "Robust dynamic bus controls considering delay disturbances and passenger demand uncertainty," Transportation Research Part B: Methodological, Elsevier, vol. 123(C), pages 88-109.
    4. Gkiotsalitis, K. & Alesiani, F., 2019. "Robust timetable optimization for bus lines subject to resource and regulatory constraints," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 128(C), pages 30-51.
    5. Wang, Zhichao & Jiang, Rui & Jiang, Yu & Gao, Ziyou & Liu, Ronghui, 2024. "Modelling bus bunching along a common line corridor considering passenger arrival time and transfer choice under stochastic travel time," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 181(C).
    6. Petit, Antoine & Lei, Chao & Ouyang, Yanfeng, 2019. "Multiline Bus Bunching Control via Vehicle Substitution," Transportation Research Part B: Methodological, Elsevier, vol. 126(C), pages 68-86.
    7. Sirmatel, Isik Ilber & Geroliminis, Nikolas, 2018. "Mixed logical dynamical modeling and hybrid model predictive control of public transport operations," Transportation Research Part B: Methodological, Elsevier, vol. 114(C), pages 325-345.
    8. Gkiotsalitis, K. & Cats, O., 2021. "At-stop control measures in public transport: Literature review and research agenda," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 145(C).
    9. Liping Ge & Stefan Voß & Lin Xie, 2022. "Robustness and disturbances in public transport," Public Transport, Springer, vol. 14(1), pages 191-261, March.
    10. Bian, Bomin & Zhu, Ning & Meng, Qiang, 2023. "Real-time cruising speed design approach for multiline bus systems," Transportation Research Part B: Methodological, Elsevier, vol. 170(C), pages 1-24.
    11. Sánchez-Martínez, G.E. & Koutsopoulos, H.N. & Wilson, N.H.M., 2016. "Real-time holding control for high-frequency transit with dynamics," Transportation Research Part B: Methodological, Elsevier, vol. 83(C), pages 1-19.
    12. Wu, Weitiao & Liu, Ronghui & Jin, Wenzhou, 2016. "Designing robust schedule coordination scheme for transit networks with safety control margins," Transportation Research Part B: Methodological, Elsevier, vol. 93(PA), pages 495-519.
    13. Hu, Sangen & Shen, Minyu & Gu, Weihua, 2023. "Impacts of bus overtaking policies on the capacity of bus stops," Transportation Research Part A: Policy and Practice, Elsevier, vol. 173(C).
    14. Zhou, Chang & Tian, Qiong & Wang, David Z.W., 2022. "A novel control strategy in mitigating bus bunching: Utilizing real-time information," Transport Policy, Elsevier, vol. 123(C), pages 1-13.
    15. Bai, Qiaowen & Ong, Ghim Ping, 2023. "Similarity-based bus services assignment with capacity constraint for staggered bus stops," Transportation Research Part E: Logistics and Transportation Review, Elsevier, vol. 179(C).
    16. Varga, Balázs & Tettamanti, Tamás & Kulcsár, Balázs, 2019. "Energy-aware predictive control for electrified bus networks," Applied Energy, Elsevier, vol. 252(C), pages 1-1.
    17. Dai, Zhuang & Liu, Xiaoyue Cathy & Chen, Zhuo & Guo, Renyong & Ma, Xiaolei, 2019. "A predictive headway-based bus-holding strategy with dynamic control point selection: A cooperative game theory approach," Transportation Research Part B: Methodological, Elsevier, vol. 125(C), pages 29-51.
    18. Xuemei Zhou & Yehan Wang & Xiangfeng Ji & Caitlin Cottrill, 2019. "Coordinated Control Strategy for Multi-Line Bus Bunching in Common Corridors," Sustainability, MDPI, vol. 11(22), pages 1-23, November.
    19. Dong Liu & Feng Xiao & Jian Luo & Fan Yang, 2023. "Deep Reinforcement Learning-Based Holding Control for Bus Bunching under Stochastic Travel Time and Demand," Sustainability, MDPI, vol. 15(14), pages 1-18, July.
    20. Federico Malucelli & Emanuele Tresoldi, 2019. "Delay and disruption management in local public transportation via real-time vehicle and crew re-scheduling: a case study," Public Transport, Springer, vol. 11(1), pages 1-25, June.

    More about this item

    Keywords

    ;
    ;
    ;
    ;
    ;

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:transe:v:200:y:2025:i:c:s1366554525001838. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/wps/find/journaldescription.cws_home/600244/description#description .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.