Abstract:Due to the complexity of the driving environment and the dynamics of the behavior of traffic participants, self-driving in dense traffic flow is very challenging. Traditional methods usually rely on predefined rules, which are difficult to adapt to various driving scenarios. Deep reinforcement learning (DRL) shows advantages over rule-based methods in complex self-driving environments, demonstrating the great potential of intelligent decision-making. However, one of the problems of DRL is the inefficiency of exploration; typically, it requires a lot of trial and error to learn the optimal policy, which leads to its slow learning rate and makes it difficult for the agent to learn well-performing decision-making policies in self-driving scenarios. Inspired by the outstanding performance of supervised learning in classification tasks, we propose a self-driving intelligent control method that combines human driving experience and adaptive sampling supervised actor-critic algorithm. Unlike traditional DRL, we modified the learning process of the policy network by combining supervised learning and DRL and adding human driving experience to the learning samples to better guide the self-driving vehicle to learn the optimal policy through human driving experience and real-time human guidance. In addition, in order to make the agent learn more efficiently, we introduced real-time human guidance in its learning process, and an adaptive balanced sampling method was designed for improving the sampling performance. We also designed the reward function in detail for different evaluation indexes such as traffic efficiency, which further guides the agent to learn the self-driving intelligent control policy in a better way. The experimental results show that the method is able to control vehicles in complex traffic environments for self-driving tasks and exhibits better performance than other DRL methods.

Multi-Objective End-to-End Self-Driving Based on Pareto-Optimal Actor-Critic Approach

Multi-objective Optimization Based Deep Reinforcement Learning for Autonomous Driving Policy

Multi-objective optimization for autonomous driving strategy based on Deep Q Network

Multi-objective Longitudinal Decision-making for Autonomous Electric Vehicle: A Entropy-constrained Reinforcement Learning Approach.

Multi-Objective Optimization of Vehicle-Following Control for Connected Electric Vehicles Based on Deep Deterministic Policy Gradient

End-to-End Urban Autonomous Driving With Safety Constraints

End-to-End Autonomous Driving Decision Method Based on Improved TD3 Algorithm in Complex Scenarios

Down-regulation of survivin by antisense oligonucleotides increases apoptosis, inhibits cytokinesis and anchorage-independent growth.

Intelligent control of self-driving vehicles based on adaptive sampling supervised actor-critic and human driving experience

Multi-Agent Soft Actor-Critic with Global Loss for Autonomous Mobility-on-Demand Fleet Control

Overcoming driving challenges in complex urban traffic: A multi-objective eco-driving strategy via safety model based reinforcement learning

Enhancing Robotic Navigation: An Evaluation of Single and Multi-Objective Reinforcement Learning Strategies

OPTIMA: Optimized Policy for Intelligent Multi-Agent Systems Enables Coordination-Aware Autonomous Vehicles

Deep Reinforcement Learning on Autonomous Driving Policy With Auxiliary Critic Network

End-to-End Autonomous Driving With Semantic Depth Cloud Mapping and Multi-Agent

Parameterized Decision-making with Multi-modal Perception for Autonomous Driving

Actor-critic objective penalty function method: an adaptive strategy for trajectory tracking in autonomous driving

A Safe and Efficient Self-evolving Algorithm for Decision-making and Control of Autonomous Driving Systems

Learning Pareto Set for Multi-Objective Continuous Robot Control

Safe Reinforcement Learning for Autonomous Vehicles through Parallel Constrained Policy Optimization