Abstract:Purpose This paper aims to realize a fully distributed multi-UAV collision detection and avoidance based on deep reinforcement learning (DRL). To deal with the problem of low sample efficiency in DRL and speed up the training. To improve the applicability and reliability of the DRL-based approach in multi-UAV control problems. Design/methodology/approach In this paper, a fully distributed collision detection and avoidance approach for multi-UAV based on DRL is proposed. A method that integrates human experience into policy training via a human experience-based adviser is proposed. The authors propose a hybrid control method which combines the learning-based policy with traditional model-based control. Extensive experiments including simulations, real flights and comparative experiments are conducted to evaluate the performance of the approach. Findings A fully distributed multi-UAV collision detection and avoidance method based on DRL is realized. The reward curve shows that the training process when integrating human experience is significantly accelerated and the mean episode reward is higher than the pure DRL method. The experimental results show that the DRL method with human experience integration has a significant improvement than the pure DRL method for multi-UAV collision detection and avoidance. Moreover, the safer flight brought by the hybrid control method has also been validated. Originality/value The fully distributed architecture is suitable for large-scale unmanned aerial vehicle (UAV) swarms and real applications. The DRL method with human experience integration has significantly accelerated the training compared to the pure DRL method. The proposed hybrid control strategy makes up for the shortcomings of two-dimensional light detection and ranging and other puzzles in applications.

UAV Head-On Situation Maneuver Generation Using Transfer-Learning-Based Deep Reinforcement Learning

Model-free Maneuvering Control of Fixed-Wing UAVs Based on Deep Reinforcement Learning

Deep Reinforcement Learning With Application to Air Confrontation Intelligent Decision-Making of Manned/Unmanned Aerial Vehicle Cooperative System

UAV Multi-Dynamic Target Interception: A Hybrid Intelligent Method Using Deep Reinforcement Learning and Fuzzy Logic

Autonomous obstacle avoidance of UAV based on deep reinforcement learning

Deep Reinforcement Learning for Flocking Motion of Multi-UAV systems: Learn from a Digital Twin

Application of Deep Reinforcement Learning in UAVs: A Review

UAV Obstacle Avoidance by Human-in-the-Loop Reinforcement in Arbitrary 3D Environment

Continuous Transfer Learning for UAV Communication-aware Trajectory Design

Human-Guided Reinforcement Learning With Sim-to-Real Transfer for Autonomous Navigation

Autonomous maneuver decision-making for a UCAV in short-range aerial combat based on an MS-DDQN algorithm

Integrating human experience in deep reinforcement learning for multi-UAV collision detection and avoidance

Subtask-masked curriculum learning for reinforcement learning with application to UAV maneuver decision-making

Maneuver Decision of UAV in Short-Range Air Combat Based on Deep Reinforcement Learning

UAV maneuver decision-making via deep reinforcement learning for short-range air combat

Collision-Avoiding Flocking With Multiple Fixed-Wing UAVs in Obstacle-Cluttered Environments: A Task-Specific Curriculum- Based MADRL Approach

Group-Based Deep Reinforcement Learning in Multi-UAV Confrontation

Multi-UAV Speed Control with Collision Avoidance and Handover-aware Cell Association: DRL with Action Branching

UAV Cooperative Air Combat Maneuvering Confrontation Based on Multi-agent Reinforcement Learning

UAV cooperative air combat maneuver decision based on multi-agent reinforcement learning

A Vision Based Deep Reinforcement Learning Algorithm for UAV Obstacle Avoidance