Abstract:Purpose Most manufacturing plants choose the easy way of completely separating human operators from robots to prevent accidents, but as a result, it dramatically affects the overall quality and speed that is expected from human–robot collaboration. It is not an easy task to ensure human safety when he/she has entered a robot’s workspace, and the unstructured nature of those working environments makes it even harder. The purpose of this paper is to propose a real-time robot collision avoidance method to alleviate this problem. Design/methodology/approach In this paper, a model is trained to learn the direct control commands from the raw depth images through self-supervised reinforcement learning algorithm. To reduce the effect of sample inefficiency and safety during initial training, a virtual reality platform is used to simulate a natural working environment and generate obstacle avoidance data for training. To ensure a smooth transfer to a real robot, the automatic domain randomization technique is used to generate randomly distributed environmental parameters through the obstacle avoidance simulation of virtual robots in the virtual environment, contributing to better performance in the natural environment. Findings The method has been tested in both simulations with a real UR3 robot for several practical applications. The results of this paper indicate that the proposed approach can effectively make the robot safety-aware and learn how to divert its trajectory to avoid accidents with humans within the workspace. Research limitations/implications The method has been tested in both simulations with a real UR3 robot in several practical applications. The results indicate that the proposed approach can effectively make the robot be aware of safety and learn how to change its trajectory to avoid accidents with persons within the workspace. Originality/value This paper provides a novel collision avoidance framework that allows robots to work alongside human operators in unstructured and complex environments. The method uses end-to-end policy training to directly extract the optimal path from the visual inputs for the scene.

Reset-Free Reinforcement Learning via Multi-State Recovery and Failure Prevention for Autonomous Robots

Failure-aware Policy Learning for Self-assessable Robotics Tasks

A Residual Meta-Reinforcement Learning Method for Training Fault-Tolerant Policies for Quadruped Robots

Meta Reinforcement Learning of Locomotion Policy for Quadruped Robots with Motor Stuck

Meta-Reinforcement Learning of Hierarchical Fault-Tolerant Controller for Multiple Leg Failures in Hexapod Robots

Reset-Free Reinforcement Learning via Multi-Task Learning: Learning Dexterous Manipulation Behaviors without Human Intervention

Dynamic Fall Recovery Control for Legged Robots via Reinforcement Learning

Learning to Recover for Safe Reinforcement Learning

When Learning Is Out of Reach, Reset: Generalization in Autonomous Visuomotor Reinforcement Learning

Safe Reinforcement Learning with Dead-Ends Avoidance and Recovery

Learning Complex Motor Skills for Legged Robot Fall Recovery

A Reinforcement Learning Approach to Automatic Error Recovery

Autonomous Algorithm for Training Autonomous Vehicles with Minimal Human Intervention

Reachability Verification Based Reliability Assessment for Deep Reinforcement Learning Controlled Robotics and Autonomous Systems

A safe reinforcement learning approach for autonomous navigation of mobile robots in dynamic environments

Automated Robot Recovery from Assumption Violations of High-Level Specifications

Recovery RL: Safe Reinforcement Learning with Learned Recovery Zones

Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning

World Models Increase Autonomy in Reinforcement Learning

Robot obstacle avoidance system using deep reinforcement learning

Accelerated Robot Learning via Human Brain Signals