Fetching the paper…
Reading the bibliography…
We present a new approach for transfer of dynamic robot control policies such as biped locomotion from simulation to real hardware.
N. Jakobi, P. Husbands, and I. Harvey, “Noise and the reality gap: The use of simulation in evolutionary robotics,” in Advances in Artificial Life , F. Morán, A. Moreno, J. J. Merelo, and P. Chacón, Eds. Berlin, Heidelberg: Springer Berlin Heidelberg, 1995, pp. 704–720
1995
Earlier work this paper cites.
N. Hansen, A. Ostermeier, and A. Gawelczyk, “On the adaptation of arbitrary normal mutation distributions in evolution strategies: The generating set adaptation.” in ICGA , 1995, pp. 57–64
1995
Earlier work this paper cites.
L. Ljung, System Identification . Boston, MA: Birkhäuser Boston, 1998, pp. 163–173. [Online]. Available: https://doi.org/10.1007/978-1-4612-1768-8_11
1998
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “ Exploration and Apprenticeship Learning in Reinforcement Learning ,” in International Conference on Machine Learning , 2005, pp. 1–8
2005
Earlier work this paper cites.
M. E. Taylor and P. Stone, “Transfer Learning for Reinforcement Learning Domains : A Survey,” vol. 10, pp. 1633–1685, 2009
2009
Earlier work this paper cites.
M. Deisenroth and C. E. Rasmussen, “Pilco: A model-based and data-efficient approach to policy search,” in Proceedings of the 28th International Conference on machine learning (ICML-11) , 2011, pp. 465–472
2011
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International Conference on Machine Learning , 2015, pp. 1889–1897
2015
Earlier work this paper cites.
S. Ha and K. Yamane, “ Reducing Hardware Experiments for Model Learning and Policy Optimization ,” IROS , 2015
2015
Earlier work this paper cites.
A. Cully, J. Clune, D. Tarapore, and J.-B. Mouret, “Robots that can adapt like animals,” Nature , vol. 521, no. 7553, p. 503, 2015
2015
Earlier work this paper cites.
J. Tan, Z. Xie, B. Boots, and C. K. Liu, “Simulation-based design of dynamic controllers for humanoid balancing,” in Intelligent Robots and Systems (IROS), 2016 IEEE/RSJ International Conference on . IEEE, 2016, pp. 2729–2736
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. Neunert, T. Boaventura, and J. Buchli, “Why off-the-shelf physics simulators fail in evaluating feedback controller performance-a case study for quadrupedal robots,” in Advances in Cooperative Robotics . World Scientific, 2017, pp. 464–472
2017
Cited alongside, same era.
W. Yu, J. Tan, C. K. Liu, and G. Turk, “Preparing for the unknown: Learning a universal policy with online system identification,” in Proceedings of Robotics: Science and Systems , Cambridge, Massachusetts, July 2017
2017
Cited alongside, same era.
A. Rajeswaran, S. Ghotra, B. Ravindran, and S. Levine, “Epopt: Learning robust neural network policies using model ensembles.” ICLR, 2017
2017
Cited alongside, same era.
A. Mandlekar, Y. Zhu, A. Garg, L. Fei-fei, and S. Savarese, “Adversarially Robust Policy Learning : Active Construction of Physically-Plausible Perturbations.” IROS, 2017
2017
Cited alongside, same era.
F. Golemo and A. A. Taïga, “Sim-to-Real Transfer with Neural-Augmented Robot Simulation,” no. CoRL, 2018
2018
Later among the works it cites.
K. Lowrey, S. Kolev, J. Dao, A. Rajeswaran, and E. Todorov, “Reinforcement learning for non-prehensile manipulation : Transfer from simulation to physical system.” SIMPAR, 2018
2018
Later among the works it cites.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . ICRA, 2018, pp. 1–8
2018
Later among the works it cites.
OpenAI, :, M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba, “Learning Dexterous In-Hand Manipulation,” ArXiv e-prints , Aug. 2018
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Gupta, C. Devin, Y. Liu, P. Abbeel, S. Levine, and C. Science, “Learning Invariant Feature Spaces To Transfer Skills With Reinforcement Learning,” no. 2008, pp. 1–14, 2017
2017
Cited alongside, same era.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning.” ICML, 2017
2017
Cited alongside, same era.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in Intelligent Robots and Systems (IROS), 2017 IEEE/RSJ International Conference on . IEEE, 2017, pp. 23–30
2017
Cited alongside, same era.
A. A. Rusu, M. Vecerik, T. Rothörl, N. Heess, R. Pascanu, and R. Hadsell, “Sim-to-real robot learning from pixels with progressive nets.” CORL, 2017
2017
Cited alongside, same era.
X. B. Peng, P. Abbeel, S. Levine, and M. van de Panne, “Deepmimic: Example-guided deep reinforcement learning of physics-based character skills,” ACM Transactions on Graphics (Proc. SIGGRAPH 2018) , 2018
2018
Cited alongside, same era.
W. Yu, G. Turk, and C. K. Liu, “Learning symmetric and low-energy locomotion,” ACM Transactions on Graphics (Proc. SIGGRAPH 2018) , vol. 37, no. 4, 2018
2018
Cited alongside, same era.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” in Proceedings of Robotics: Science and Systems , Pittsburgh, Pennsylvania, June 2018
2018
Cited alongside, same era.
H.-w. Park, K. Sreenath, J. W. Hurst, and J. W. Grizzle, “Identification and Dynamic Model of a Bipedal Robot With a Cable-Differential-Based Compliant Drivetrain,” pp. 1–17
Cited in the paper.
T. Chen, A. Murali, and A. Gupta, “Hardware conditioned policies for multi-robot transfer learning.” NIPS, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
J. Lee, M. X. Grey, S. Ha, T. Kunz, S. Jain, Y. Ye, S. S. Srinivasa, M. Stilman, and C. K. Liu, “Dart: Dynamic animation and robotics toolkit,” The Journal of Open Source Software , vol. 3, no. 22, p. 500, 2018
2018
Later among the works it cites.
W. Yu, C. K. Liu, and G. Turk, “Policy transfer with strategy optimization,” in International Conference on Learning Representations , 2019. [Online]. Available: https://openreview.net/forum?id=H1g6osRcFQ
2019
Closest in time.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, p. eaau5872, 2019
2019
Closest in time.
Y. Chebotar, A. Handa, V. Makoviychuk, M. Macklin, J. Issac, N. Ratliff, and D. Fox, “Closing the Sim-to-Real Loop : Adapting Simulation Randomization with Real World Experience.” ICRA, 2019
2019
Closest in time.