DroneDeep RL (DDR): A Traffic Congestion Control Strategy Using Prioritization LLM Agent and Circular Deep Q-Network †
Abstract
1. Introduction
2. Proposed Methodology
2.1. System Overview
2.2. LLM Agent-Based Congestion-Aware Prioritization
2.3. Proposed DRL Controller
3. Simulation Result
3.1. Average Waiting Time Analysis While Prioritizing Areas
3.2. Cumulative Rewards per Episode
3.3. Average Waiting Time Delay
4. Conclusions
Author Contributions
Funding
Institutional Review Board Statement
Informed Consent Statement
Data Availability Statement
Conflicts of Interest
References
- Kumar, S.; Vishal; Sharma, P.; Pal, N. Object tracking and counting in a zone using YOLOv4, DeepSORT and TensorFlow. In Proceedings of the IEEE International Conference on Artificial Intelligence and Smart Systems (ICAIS), Coimbatore, India, 25–27 March 2021; pp. 1017–1022. [Google Scholar] [CrossRef] [Scilit]
- Asha, C.S.; Narasimhadhan, A.V. Vehicle counting for traffic management system using YOLO and correlation filter. In Proceedings of the IEEE International Conference Communication and Signal Processing (ICCSP), Tamilnadu, India, 3–5 April 2018; pp. 1–5. [Google Scholar]
- Chaudhuri, A. Smart Traffic Management of Vehicles Using Faster R-CNN Based Deep Learning Method. Sci. Rep. 2024, 14, 10357. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Mahmud, D.; Hajmohamed, H.; Almentheri, S.; Alqaydi, S.; Aldhaheri, L.; Khalil, R.A.; Saeed, N. Integrating LLMs with ITS: Recent advances, potentials, challenges, and future directions. arXiv 2025, arXiv:2501.04437v1. [Google Scholar] [CrossRef] [Scilit]
- Al-Fuqaha, A.; Guizani, M.; Mohammadi, M.; Aledhari, M.; Ayyash, M. Internet of Things: A survey on enabling technologies, protocols, and applications. IEEE Commun. Surv. Tutor. 2015, 17, 2347–2376. [Google Scholar] [CrossRef] [Scilit]
- Mohammed, A.K.; Anwer, R.; Hussain, F. Drone-Based Real-Time Traffic Monitoring System Using Computer Vision. IEEE Access 2022, 10, 84562–84573. [Google Scholar]
- Lin, T.; Chen, L.; Li, Q. Lightweight Object Detection Network for UAV Traffic Monitoring. Sensors 2022, 22, 7321–7335. [Google Scholar]
- Wang, C.; Zhang, X.; Xu, J. YOLOv11: Enhanced Lightweight Object Detection for Edge Devices. arXiv 2025, arXiv:2503.01456. [Google Scholar]
- Li, L.; Lv, Y.; Wang, F.-Y. Traffic Signal Timing via Deep Reinforcement Learning. IEEE/CAA J. Autom. Sin. 2016, 3, 247–254. [Google Scholar] [CrossRef] [Scilit]
- Kővári, B.; Tamás, T.; Bécsi, T. Deep Reinforcement Learning-Based Approach for Traffic Signal Control. Procedia Comput. Sci. 2022, 62, 278–285. [Google Scholar]
- Chu, Y.; Wang, X.; Gao, Y. Multi-Agent deep reinforcement learning for large-scale traffic signal control. IEEE Trans. Intell. Transp. Syst. 2022, 23, 4482–4496. [Google Scholar] [CrossRef] [Scilit]
- Wei, Z.; Chen, C.; Zheng, K.; Wang, X. IntelliLight: A Reinforcement Learning Approach for Intelligent Traffic Light Control. In Proceedings of the ACM SIGKDD International Conference Knowledge Discovery & Data Mining, London, UK, 19–23 August 2018; pp. 2496–2505. [Google Scholar]
- Zheng, Y.; Luo, J.; Gao, H.; Zhou, Y.; Li, K. Pri-DDQN: Learning Adaptive Traffic Signal Control Strategy. Complex Intell. Syst. 2025, 11, 4. [Google Scholar] [CrossRef] [Scilit]
- Kwesiga, D.K.; Guin, A.; Hunter, M. Adaptive Traffic Signal Control based on Multi-Agent Reinforcement Learning: Case Study on a Simulated Real-World Corridor. arXiv 2025, arXiv:2503.02189. [Google Scholar]
- Bohra, A.R.; Selvi, T. Reinforcement Learning for Adaptive Traffic Signal Control Using Deep Q-Networks. In Proceedings of the 2025 International Conference on Sustainable Energy Technologies and Computational Intelligence (SETCOM), Gandhinagar, India, 21–23 February 2025. [Google Scholar] [CrossRef] [Scilit]
- Alif, M.A.R. YOLOv11 for vehicle detection: Advancements, performance, and applications in intelligent transportation systems. arXiv 2024, arXiv:2410.22898. Available online: https://arxiv.org/abs/2410.22898 (accessed on 30 October 2024). [CrossRef] [Scilit]
- Talbi, D.; Boukhtouta, A.; Chraibi, M.; Tembine, H. Integrating Reinforcement Learning into M/M/1/K Retry Queuing for 6G Networks. Preprints 2025. Available online: https://pmc.ncbi.nlm.nih.gov/ (accessed on 9 April 2026).






Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
© 2026 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license.
Share and Cite
Hasan, M.M.; Siddika, A.; Khushi, M.A.; Sultan, S.M.; Alam, T.; Arman, S.H. DroneDeep RL (DDR): A Traffic Congestion Control Strategy Using Prioritization LLM Agent and Circular Deep Q-Network. Eng. Proc. 2026, 129, 30. https://doi.org/10.3390/engproc2026129030
Hasan MM, Siddika A, Khushi MA, Sultan SM, Alam T, Arman SH. DroneDeep RL (DDR): A Traffic Congestion Control Strategy Using Prioritization LLM Agent and Circular Deep Q-Network. Engineering Proceedings. 2026; 129(1):30. https://doi.org/10.3390/engproc2026129030
Chicago/Turabian StyleHasan, Md. Mujahid, Afsana Siddika, Maria Akter Khushi, Salman Md Sultan, Tahira Alam, and Shajedul Hasan Arman. 2026. "DroneDeep RL (DDR): A Traffic Congestion Control Strategy Using Prioritization LLM Agent and Circular Deep Q-Network" Engineering Proceedings 129, no. 1: 30. https://doi.org/10.3390/engproc2026129030
APA StyleHasan, M. M., Siddika, A., Khushi, M. A., Sultan, S. M., Alam, T., & Arman, S. H. (2026). DroneDeep RL (DDR): A Traffic Congestion Control Strategy Using Prioritization LLM Agent and Circular Deep Q-Network. Engineering Proceedings, 129(1), 30. https://doi.org/10.3390/engproc2026129030
