Extending the OpenAI Gym for robotics: a toolkit for reinforcement learning using ROS and Gazebo
This paper presents an extension of the OpenAI Gym for robotics using the Robot Operating System (ROS) and the Gazebo simulator. The content discusses the software architecture proposed and the results obtained by using two Reinforcement Learning techniques: Q-Learning and Sarsa. Ultimately, the output of this work presents a benchmarking system for robotics that allows different techniques and algorithms to be compared using the same virtual conditions.
Paper
References (11)
01OpenAI GymGreg Brockman, Vicki Cheung, Ludwig Pettersson et al.2016 · arXiv.org · 5.6k citations In Library
07Gazebo-3d multiple robot simulator with dynamics2006 · Gazebo-3d multiple robot simulator with dynamics
09Qlearning”. In: Machine learning1992 · Erle Robotics
10Reinforcement Learning: Q-Learning vs Sarsa. [Online; accessed 7-August-2016]Reinforcement Learning: Q-Learning vs Sarsa. [Online; accessed 7-August-2016]
11erlerobotics.com