Fetching the paper…
Reading the bibliography…
Computer simulation provides an automatic and safe way for training robotic control policies to achieve complex tasks such as locomotion.
On the adaptation of arbitrary normal mutation distributions in evolution strategies: The generating set adaptation
Nikolaus Hansen, Andreas Ostermeier, and Andreas Gawelczyk · 1995
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Exploration and Apprenticeship Learning in Reinforcement Learning
Pieter Abbeel and Andrew Y. Ng · 2005
Earlier work this paper cites.
Crossing the reality gap in evolutionary robotics by promoting transferable controllers
Sylvain Koos, Jean-Baptiste Mouret, and Stéphane Doncieux · 2010
Earlier work this paper cites.
Pilco: A model-based and data-efficient approach to policy search
Marc Deisenroth and Carl E Rasmussen · 2011
Earlier work this paper cites.
Controlling physics-based characters using soft contacts
Sumit Jain and C Karen Liu · 2011
Earlier work this paper cites.
Leveraging multiple simulators for crossing the reality gap
Adrian Boeing and Thomas Bräunl · 2012
Earlier work this paper cites.
Bruno Da Silva, George Konidaris, and Andrew Barto · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Learning environmental calibration actions for policy self-evolution
Chao Zhang, Yang Yu, and Zhi-Hua Zhou · 2012
Earlier work this paper cites.
Learning compact parameterized skills with a single regression
Freek Stulp, Gennaro Raiola, Antoine Hoarau, Serena Ivaldi, and Olivier Sigaud · 2013
Earlier work this paper cites.
Robots that can adapt like animals
Antoine Cully, Jeff Clune, Danesh Tarapore, and Jean-Baptiste Mouret · 2015
Earlier work this paper cites.
Reducing Hardware Experiments for Model Learning and Policy Optimization
Sehoon Ha and Katsu Yamane · 2015
Earlier work this paper cites.
Trust region policy optimization
John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz · 2015
Cited alongside, same era.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Cited alongside, same era.
Pydart2, 2016
Sehoon Ha · 2016
Cited alongside, same era.
Epopt: Learning robust neural network policies using model ensembles
Aravind Rajeswaran, Sarvjeet Ghotra, Balaraman Ravindran, and Sergey Levine · 2016
Cited alongside, same era.
Sim-to-real robot learning from pixels with progressive nets
Andrei A Rusu, Matej Vecerik, Thomas Rothörl, Nicolas Heess, Razvan Pascanu, and Raia Hadsell · 2016
Cited alongside, same era.
Domain randomization for transferring deep neural networks from simulation to the real world
Josh Tobin, Rachel Fong, Alex Ray, Jonas Schneider, Wojciech Zaremba, and Pieter Abbeel · 2017
Later among the works it cites.
Mutual alignment transfer learning
Markus Wulfmeier, Ingmar Posner, and Pieter Abbeel · 2017
Later among the works it cites.
Dartenv, 2017
Wenhao Yu and C. Karen Liu · 2017
Later among the works it cites.
Preparing for the unknown: Learning a universal policy with online system identification
Wenhao Yu, Jie Tan, C. Karen Liu, and Greg Turk · 2017
Later among the works it cites.
Hardware conditioned policies for multi-robot transfer learning
Tao Chen, Adithyavairavan Murali, and Abhinav Gupta · 2018
Closest in time.
Learning to dress: Synthesizing human dressing motion via deep reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Simulation-based design of dynamic controllers for humanoid balancing
Jie Tan, Zhaoming Xie, Byron Boots, and C Karen Liu · 2016
Cited alongside, same era.
Pybullet, a python module for physics simulation in robotics, games and machine learning., 2016-2017
Erwin Coumans and Yunfei Bai · 2017
Cited alongside, same era.
Openai baselines
Prafulla Dhariwal, Christopher Hesse, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, and Yuhuai Wu · 2017
Cited alongside, same era.
Emergence of locomotion behaviours in rich environments
Nicolas Heess, Srinivasan Sriram, Jay Lemmon, Josh Merel, Greg Wayne, Yuval Tassa, Tom Erez, Ziyu Wang, Ali Eslami, Martin Riedmiller, et al · 2017
Cited alongside, same era.
Why off-the-shelf physics simulators fail in evaluating feedback controller performance-a case study for quadrupedal robots
Michael Neunert, Thiago Boaventura, and Jonas Buchli · 2017
Cited alongside, same era.
Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning
Xue Bin Peng, Glen Berseth, KangKang Yin, and Michiel Van De Panne · 2017
Cited alongside, same era.
Robust adversarial reinforcement learning
Lerrel Pinto, James Davidson, Rahul Sukthankar, and Abhinav Gupta · 2017
Cited alongside, same era.
Alexander Clegg, Wenhao Yu, Jie Tan, C. Karen Liu, and Greg Turk · 2018
Closest in time.
Diversity is all you need: Learning skills without a reward function
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz, and Sergey Levine · 2018
Closest in time.
Dart: Dynamic animation and robotics toolkit
Jeongseok Lee, Michael X Grey, Sehoon Ha, Tobias Kunz, Sumit Jain, Yuting Ye, Siddhartha S Srinivasa, Mike Stilman, and C Karen Liu · 2018
Closest in time.
Learning Dexterous In-Hand Manipulation
OpenAI, :, M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba · 2018
Closest in time.
Sim-to-real transfer of robotic control with dynamics randomization
Xue Bin Peng, Marcin Andrychowicz, Wojciech Zaremba, and Pieter Abbeel · 2018
Closest in time.
Sim-to-real: Learning agile locomotion for quadruped robots
Jie Tan, Tingnan Zhang, Erwin Coumans, Atil Iscen, Yunfei Bai, Danijar Hafner, Steven Bohez, and Vincent Vanhoucke · 2018
Closest in time.
Learning symmetric and low-energy locomotion
Wenhao Yu, Greg Turk, and C. Karen Liu · 2018
Closest in time.