Abstract:Path planning is one of the research hotspots for outdoor mobile robots. This paper addresses the issues of slow convergence and low accuracy in the Double Deep Q Network (DDQN) method in environments with many obstacles in the context of deep reinforcement learning. A new algorithm, Improve Double Deep Q Network (IDDQN), is proposed, which utilizes second-order temporal difference methods and a binary tree data structure to improve the DDQN method. The improved method evaluates the actions of the current robot using second-order temporal difference methods and employs a binary tree structure to store the results obtained from these methods, replacing the traditional experience pool structure. The environment is constructed using a grid method, programmed in the Python language, with two two-dimensional grid maps created for simple and complex environments. DDQN and four related deep reinforcement learning methods, such as Multi-step updates and Experience Classification Double Deep Q Network (ECMS-DDQN), are compared through simulation experiments with the IDDQN method. Simulation results indicate that the IDDQN method improves various path planning metrics compared to the DDQN method and other relevant reinforcement learning methods. In the simple environment, IDDQN method exhibits a 26.89% improvement in step convergence time, a 22.58% improvement in reward convergence time, and a 10.30% improvement in average reward value after convergence compared to the original DDQN algorithm. It also outperforms other simulated methods in the simple environment, although the difference is not significant. In the complex environment, the IDDQN method avoids falling into local optima compared to other methods, demonstrating the accuracy of its strategy in complex environments. Other methods show artificially high average reward values after converging in local optima, lacking reference value. In the complex environment, IDDQN method exhibits a 33.22% improvement in step convergence time and a 25.47% improvement in reward convergence time compared to the original DDQN algorithm, clearly surpassing other participating simulated methods. The data above indicate that the IDDQN method improves both convergence speed and accuracy compared to the DDQN method and the relevant improvement methods simulated in this paper. Particularly in environments with many obstacles, the performance improvement is evident, allowing for effective path planning in such environments.

Research on Dynamic Path Planning of Wheeled Robot Based on Deep Reinforcement Learning on the Slope Ground

Path Planning of Autonomous Mobile Robot in Comprehensive Unknown Environment Using Deep Reinforcement Learning

Improved Robot Path Planning Method Based on Deep Reinforcement Learning

SLP-Improved DDPG Path-Planning Algorithm for Mobile Robot in Large-Scale Dynamic Environment

Research on mobile robot path planning in complex environment based on DRQN algorithm

Dynamic Path Planning of Unknown Environment Based on Deep Reinforcement Learning

Path planning for outdoor mobile robots based on IDDQN (October 2023)

Research on Dynamic Path Planning of Mobile Robot Based on Improved DDPG Algorithm

Path Planning Method of Mobile Robot Using Improved Deep Reinforcement Learning

Path Planning for Autonomous Vehicles in Unknown Dynamic Environment Based on Deep Reinforcement Learning

An Improved Algorithm of Robot Path Planning in Complex Environment Based on Double DQN

Dynamic Path Planning for Mobile Robots with Deep Reinforcement Learning

Path planning of mobile robot based on improved TD3 algorithm in dynamic environment

Path Following for Autonomous Ground Vehicle Using DDPG Algorithm: A Reinforcement Learning Approach

Efficient Path Planning for Mobile Robot Based on Deep Deterministic Policy Gradient

A Mapless Local Path Planning Approach Using Deep Reinforcement Learning Framework

Mapless Path Planning for Mobile Robot Based on Improved Deep Deterministic Policy Gradient Algorithm

A Path Planning Algorithm Based on Deep Reinforcement Learning for Mobile Robots in Unknown Environment

Deep Reinforcement Learning for Mobile Robot Path Planning

High Maneuverability Control of Single-track Two-wheeled Robot in Narrow Terrain Based on Reinforcement Learning.

Deep Reinforcement Learning for Indoor Mobile Robot Path Planning