Online longitudinal trajectory planning for connected and autonomous vehicles in mixed traffic flow with deep reinforcement learning approach

Yanqiu Cheng,Xianbiao Hu,Kuanmin Chen,Xinlian Yu,Yulong Luo

doi:10.1080/15472450.2022.2046472

Abstract

This manuscript presents an Adam optimization-based Deep Reinforcement Learning model for Mixed Traffic Flow control (ADRL-MTF), so as to guide Connected and Autonomous vehicle’s (CAV) longitudinal trajectory on a typical urban roadway with signal-controlled intersections. Two improvements are made when compared with prior literatures. First, the common assumptions to simplify the problem solving, such as dividing a vehicle trajectory into several segments with constant acceleration/deceleration, are avoided, to improve the modeling realism. Second, built on the efficient Adam Optimization and Deep Q-Learning, the proposed model avoids the enumeration of states and actions, and is computational efficient and suitable for real time applications. The mixed traffic flow dynamic is firstly formulated as a finite Markov decision process (MDP) model. Due to the discretization of time, space and speed, this MDP model becomes high-dimensional in state, and is very challenging to solve. We then propose a temporal difference-based deep reinforcement learning approach, with -greedy for exploration-exploitation balance. Two neural networks are developed to replace the traditional Q function and generate the targets in the Q-learning update. These two neural networks are trained by the Adam optimization algorithm, which extends stochastic gradient descent and considers the second moments of the gradients, and is thus highly computational efficient and has lower memory requirements. The proposed model is shown to reduce fuel consumption by 7.8%, which outperforms a prior benchmark model based on Monte Carlo Tree Search. The model’s runtime efficiency and stability are tested, and the sensitivity analysis is also performed.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Online longitudinal trajectory planning for connected and autonomous vehicles in mixed traffic flow with deep reinforcement learning approach

Abstract

Talk to us

Similar Papers

More From: Journal of Intelligent Transportation Systems

Lead the way for us

Journal: Journal of Intelligent Transportation Systems	Publication Date: Feb 24, 2022
Citations: 15

Similar Papers

AlphaTruss: Monte Carlo Tree Search for Optimal Truss Layout Design
Ruifeng Luo ... Xianzhong Zhao
Buildings | VOL. 12
Ruifeng Luo, et. al.Ruifeng Luo ... Xianzhong Zhao
11 May 2022
Buildings | VOL. 12

UAV-Assisted Wireless Energy and Data Transfer With Deep Reinforcement Learning
Zehui Xiong ... Jiawen Kang
IEEE Transactions on Cognitive Communications and Networking | VOL. 7
Zehui Xiong, et. al.Zehui Xiong ... Jiawen Kang
30 Sep 2020
IEEE Transactions on Cognitive Communications and Networking | VOL. 7

Markov decision process model for patient admission decision at an emergency department under a surge demand
Hyun-Rok Lee ... Taesik Lee
Flexible Services and Manufacturing Journal | VOL. 30
Hyun-Rok Lee, et. al.Hyun-Rok Lee ... Taesik Lee
08 Feb 2017
Flexible Services and Manufacturing Journal | VOL. 30

Optimal production ramp‐up in the smartphone manufacturing industry
Lu Wang ... Changjing Hong
Naval Research Logistics (NRL) | VOL. 67
Lu Wang, et. al.Lu Wang ... Changjing Hong
10 Jan 2020
Naval Research Logistics (NRL) | VOL. 67

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Online longitudinal trajectory planning for connected and autonomous vehicles in mixed traffic flow with deep reinforcement learning approach

Abstract

Talk to us

Similar Papers

More From: Journal of Intelligent Transportation Systems