Intelligent control of self-driving vehicles based on adaptive sampling supervised actor-critic and human driving experience-Reference-Cited by-同舟云学术

Intelligent control of self-driving vehicles based on adaptive sampling supervised actor-critic and human driving experience

Published:2024 Issue:5 Volume:21 Page:6077-6096
ISSN:1551-0018
Container-title:Mathematical Biosciences and Engineering
language:
Short-container-title:MBE

Author:

Zhang Jin¹,Ma Nan²,Wu Zhixuan³,Wang Cheng¹,Yao Yongqiang⁴

Affiliation:

1. Beijing Key Laboratory of Information Service Engineering, Beijing Union University, Beijing 100101, China

2. Faculty of Information Technology, Beijing University of Technology, Beijing 100124, China

3. Beijing University of Posts and Telecommunications, Beijing 100876, China

4. Beijing Shuncheng High Technology Corporation, Beijing 102206, China

Abstract

<abstract><p>Due to the complexity of the driving environment and the dynamics of the behavior of traffic participants, self-driving in dense traffic flow is very challenging. Traditional methods usually rely on predefined rules, which are difficult to adapt to various driving scenarios. Deep reinforcement learning (DRL) shows advantages over rule-based methods in complex self-driving environments, demonstrating the great potential of intelligent decision-making. However, one of the problems of DRL is the inefficiency of exploration; typically, it requires a lot of trial and error to learn the optimal policy, which leads to its slow learning rate and makes it difficult for the agent to learn well-performing decision-making policies in self-driving scenarios. Inspired by the outstanding performance of supervised learning in classification tasks, we propose a self-driving intelligent control method that combines human driving experience and adaptive sampling supervised actor-critic algorithm. Unlike traditional DRL, we modified the learning process of the policy network by combining supervised learning and DRL and adding human driving experience to the learning samples to better guide the self-driving vehicle to learn the optimal policy through human driving experience and real-time human guidance. In addition, in order to make the agent learn more efficiently, we introduced real-time human guidance in its learning process, and an adaptive balanced sampling method was designed for improving the sampling performance. We also designed the reward function in detail for different evaluation indexes such as traffic efficiency, which further guides the agent to learn the self-driving intelligent control policy in a better way. The experimental results show that the method is able to control vehicles in complex traffic environments for self-driving tasks and exhibits better performance than other DRL methods.</p></abstract>

Publisher

American Institute of Mathematical Sciences (AIMS)

Reference39 articles.

1. B. R. Kiran, I. Sobh, V. Talpaert, P. Mannion, A. A. A. Sallab, S. Yogamani, et al., Deep reinforcement learning for autonomous driving: A survey, IEEE Trans. Intell. Transp. Syst., 23 (2022), 4909–4926. https://doi.org/10.1109/TITS.2021.3054625

2. J. Chen, B. Yuan, M. Tomizuka, Model-free deep reinforcement learning for urban autonomous driving, in 2019 IEEE intelligent transportation systems conference (ITSC), (2019), 2765–2771. https://doi.org/10.1109/ITSC.2019.8917306

3. M. Panzer, B. Bender, Deep reinforcement learning in production systems: a systematic literature review, Int. J. Prod. Res., 60 (2022), 4316–4341. https://doi.org/10.1080/00207543.2021.1973138

4. N. Ma, Y. Gao, J. Li, D. Li, Interactive cognition in self-driving, Sci. Sin. Inf., 48 (2018), 1083–1096. https://doi.org/10.1360/N112018-00028

5. H. Shi, D. Chen, N. Zheng, X. Wang, Y. Zhou, B. Ran, A deep reinforcement learning based distributed control strategy for connected automated vehicles in mixed traffic platoon, Transp. Res. Part C: Emerging Technol., 148 (2023), 104019. https://doi.org/10.1016/j.trc.2023.104019