Fetching the paper…
Reading the bibliography…
This paper presents an extension of the OpenAI Gym for robotics using the Robot Operating System (ROS) and the Gazebo simulator.
“Q-learning”
Christopher˜JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
“On-line Q-learning using connectionist systems”
Gavin˜A Rummery and Mahesan Niranjan · 1994
Earlier work this paper cites.
“Gazebo-3d multiple robot simulator with dynamics”
Nathan Koenig and Andrew Howard · 2006
Earlier work this paper cites.
“ROS: an open-source Robot Operating System”
Morgan Quigley et al · 2009
Earlier work this paper cites.
“Reinforcement learning in robotics: A survey”
Jens Kober, J˜Andrew Bagnell and Jan Peters · 2013
Cited alongside, same era.
“Reinforcement learning in robotics: Applications and real-world challenges”
Petar Kormushev, Sylvain Calinon and Darwin˜G Caldwell · 2013
Cited alongside, same era.
URL: https://studywolf.wordpress.com/2013/07/01/reinforcement-learning-sarsaverb-vs-q-learning/
“Reinforcement Learning: Sarsa vs Qlearn” [Online; accessed 7-August-2016] · 2013
Cited alongside, same era.
Greg Brockman et al · 2016
Closest in time.
URL: http://blog.deeprobotics.es/robots,/ai,/deep/learning,/rl,/reinforcemenverbt/learning/2016/07/06/rl-intro/
“Reinforcement learning in robotics” [Online; accessed 6-August-2016] · 2016
Closest in time.
URL: http://www.cse.unsw.edu.au/~cs9417ml/RL1/algorithms.html
“Reinforcement Learning: Q-Learning vs Sarsa” [Online; accessed 7-August-2016] · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…