Abstract:Deep Reinforcement Learning (DRL) is a quickly evolving research field rooted in operations research and behavioural psychology, with potential applications extending across various domains, including robotics. This thesis delineates the background of modern Reinforcement Learning (RL), starting with the framework constituted by the Markov decision processes, Markov properties, goals and rewards, agent-environment interactions, and policies. We explain the main types of algorithms commonly used in RL, including value-based, policy gradient, and actor-critic methods, with a special emphasis on DQN, A2C and PPO. We then give a short literature review on some widely adopted frameworks for implementing RL algorithms and environments. Subsequently, we present Bidimensional Gripper Environment (BGE), a virtual simulator based on the Pymunk physics engine we developed to analyse top-down bidimensional object manipulation. The methodology section frames our agent-environment interaction as a Markov decision process, such that we can apply our RL algorithms. We list various goal formulation strategies, including reward shaping and curriculum learning. We also employ different steps of observation preprocessing to reduce the computational workload required. In the experimental phase, we run through a series of scenarios of increasing difficulty. We start with a simple static scenario and then gradually increase the amount of stochasticity. Whenever the agents show difficulty in learning, we counteract by increasing the degree of reward shaping and curriculum learning. These experiments demonstrate the substantial limitations and pitfalls of model-free algorithms under changing dynamics. In conclusion, we present a summary of our findings and remarks. We then outline potential future work to improve our methodology and possibly expand to real-world systems.

Intrinsic Motivation Driven Intuitive Physics Learning using Deep Reinforcement Learning with Intrinsic Reward Normalization

Intuitive physics learning in a deep-learning model inspired by developmental psychology

Learning Intuitive Physics and One-Shot Imitation Using State-Action-Prediction Self-Organizing Maps

Physics-Regulated Deep Reinforcement Learning: Invariant Embeddings

Learning to Play Video Games with Intuitive Physics Priors

3D-IntPhys: Towards More Generalized 3D-grounded Visual Intuitive Physics under Challenging Scenes

Physics-Informed Model-Based Reinforcement Learning

DeepMimic: Example-Guided Deep Reinforcement Learning of Physics-Based Character Skills

Physics-Guided Hierarchical Reward Mechanism for Learning-Based Robotic Grasping

Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation

Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation

Latent Intuitive Physics: Learning to Transfer Hidden Physics from A 3D Video

Curiosity-driven Intuitive Physics Learning

Combining physics and deep learning to learn continuous-time dynamics models

Strategy and Skill Learning for Physics-based Table Tennis Animation

Architecting and Visualizing Deep Reinforcement Learning Models

Deep Reinforcement Learning for 2D Physics-Based Object Manipulation in Clutter

Image-Based Deep Reinforcement Learning with Intrinsically Motivated Stimuli: On the Execution of Complex Robotic Tasks

Learning To Estimate Regions Of Attraction Of Autonomous Dynamical Systems Using Physics-Informed Neural Networks

Physical Deep Reinforcement Learning Towards Safety Guarantee

PhyPlan: Generalizable and Rapid Physical Task Planning with Physics Informed Skill Networks for Robot Manipulators