Abstract:Proceedings of the Institution of Mechanical Engineers, Part D: Journal of Automobile Engineering, Ahead of Print. Present the DDPGwP (DDPG with Pretraining) model, grounded in the framework of deep reinforcement learning, designed for autonomous driving decision-making. The model incorporates imitation learning by utilizing expert experience for supervised learning during initial training and weight preservation. A novel loss function is devised, enabling the expert experience to jointly guide the Actor network's update alongside the Critic network while also participating in the Critic network's updates. This approach allows imitation learning to dominate the early stages of training, with reinforcement learning taking the lead in later stages. Employing experience replay buffer separation techniques, we categorize and store collected superior, ordinary, and expert experiences. We select sensor inputs from the TORCS (The Open Racing Car Simulator) simulation platform and conduct experimental validation, comparing the results with the original DDPG, A2C, and PPO algorithms. Experimental outcomes reveal that incorporating imitation learning significantly accelerates early-stage training, reduces blind trial-and-error during initial exploration, and enhances algorithm stability and safety. The experience replay buffer separation technique improves sampling efficiency and mitigates algorithm overfitting. In addition to expediting algorithm training rates, our approach enables the simulated vehicle to learn superior strategies, garnering higher reward values. This demonstrates the superior stability, safety, and policy-making capabilities of the proposed algorithm, as well as accelerated network convergence.

Dyna-PPO Reinforcement Learning with Gaussian Process for the Continuous Action Decision-Making in Autonomous Driving.

Decision-making for Autonomous Vehicles on Highway: Deep Reinforcement Learning with Continuous Action Horizon

Path Following for Autonomous Ground Vehicle Using DDPG Algorithm: A Reinforcement Learning Approach

End-To-End Autonomous Driving Decision Based On Deep Reinforcement Learning

An Improved Proximal Policy Optimization Algorithm for Autonomous Driving Decision-Making

Self-Driving Via Improved DDPG Algorithm

Research on Autonomous Driving Decision-making Strategies based Deep Reinforcement Learning

Learn to Make Decision with Small Data for Autonomous Driving: Deep Gaussian Process and Feedback Control

Gaussian Process Based Deep Dyna-Q Approach for Dialogue Policy Learning.

Efficient Reinforcement Learning in Continuous State and Action Spaces with Dyna and Policy Approximation.

Deep Deterministic Policy Gradient Algorithm Based on Convolutional Block Attention for Autonomous Driving.

Integration of Decision-Making and Motion Planning for Autonomous Driving Based on Double-Layer Reinforcement Learning Framework

A decision-making of autonomous driving method based on DDPG with pretraining

An Automated Driving Strategy Generating Method Based on WGAIL–DDPG

An Integrated Framework of Lateral and Longitudinal Behavior Decision-Making for Autonomous Driving Using Reinforcement Learning

Large Language Model guided Deep Reinforcement Learning for Decision Making in Autonomous Driving

Path tracking control based on Deep reinforcement learning in Autonomous driving

Deep Reinforcement Learning on Autonomous Driving Policy With Auxiliary Critic Network

Continuous Reinforcement Learning From Human Demonstrations With Integrated Experience Replay For Autonomous Driving

Multi-policy Soft Actor-Critic Reinforcement Learning for Autonomous Racing

Deep reinforcement learning for autonomous driving in uncontrolled intersections of Indian roads