Fetching the paper…
Reading the bibliography…
Legged robots have unparalleled mobility on unstructured terrains.
M. McCloskey and N. Cohen, “Catastrophic interference in connectionist networks: The sequential learning problem,” Psychology of Learning and Motivation - Advances in Research and Theory , vol. 24, no. C, pp. 109–165, Jan. 1989
1989
Earlier work this paper cites.
R. Caruana, “Multitask learning: A knowledge-based source of inductive bias,” in Proceedings of the Tenth International Conference on International Conference on Machine Learning , ser. ICML’93. San Francisco, CA, USA: Morgan Kaufmann Publishers Inc., 1993, p. 41–48
1993
Earlier work this paper cites.
J. S. Matthis and B. R. Fajen, “Visual control of foot placement when walking over complex terrain,” J Exp Psychol Hum Percept Perform , vol. 40, no. 1, pp. 106–115, Feb 2014
2014
Earlier work this paper cites.
D. Kim, Y. Zhao, G. Thomas, B. R. Fernandez, and L. Sentis, “Stabilizing series-elastic point-foot bipeds using whole-body operational space control,” IEEE Transactions on Robotics , vol. 32, no. 6, pp. 1362–1379, 2016
2016
Earlier work this paper cites.
N. Heess, D. TB, S. Sriram, J. Lemmon, J. Merel, G. Wayne, Y. Tassa, T. Erez, Z. Wang, S. M. A. Eslami, M. Riedmiller, and D. Silver, “Emergence of locomotion behaviours in rich environments,” 2017
2017
Earlier work this paper cites.
X. B. Peng, G. Berseth, K. Yin, and M. van de Panne, “Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning,” ACM Transactions on Graphics (Proc. SIGGRAPH 2017) , vol. 36, no. 4, 2017
2017
Earlier work this paper cites.
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov, “Proximal policy optimization algorithms,” 2017
2017
Earlier work this paper cites.
C. D. Bellicoso, M. Bjelonic, L. Wellhausen, K. Holtmann, F. Günther, M. Tranzatto, P. Fankhauser, and M. Hutter, “Advances in real-world applications for legged robots,” Journal of Field Robotics , vol. 35, no. 8, pp. 1311–1326, 2018. [Online]. Available: https://onlinelibrary.wiley.com/doi/abs/10.1002/rob.21839
2018
Earlier work this paper cites.
A. Iscen, K. Caluwaerts, J. Tan, T. Zhang, E. Coumans, V. Sindhwani, and V. Vanhoucke, “Policies modulating trajectory generators,” in Journal of Machine Learning Research , ser. Proceedings of Machine Learning Research, A. Billard, A. Dragan, J. Peters, and J. Morimoto, Eds., vol. 87. PMLR, 29–31 Oct 2018, pp. 916–926. [Online]. Available: http://proceedings.mlr.press/v87/iscen18a.html
2018
Earlier work this paper cites.
F. Xia, A. R. Zamir, Z.-Y. He, A. Sax, J. Malik, and S. Savarese, “Gibson env: real-world perception for embodied agents,” in Computer Vision and Pattern Recognition (CVPR), 2018 IEEE Conference on . IEEE, 2018
2018
Earlier work this paper cites.
A. W. Winkler, C. D. Bellicoso, M. Hutter, and J. Buchli, “Gait and trajectory optimization for legged systems through phase-based end-effector parameterization,” IEEE Robotics and Automation Letters , vol. 3, no. 3, pp. 1560–1567, 2018
2018
Earlier work this paper cites.
G. Bledt, M. Powell, B. Katz, J. Carlo, P. Wensing, and S. Kim, “Mit cheetah 3: Design and control of a robust, dynamic quadruped robot,” 10 2018
2018
Cited alongside, same era.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Cited alongside, same era.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” 2018
2018
Cited alongside, same era.
M. Hessel, H. Soyer, L. Espeholt, W. Czarnecki, S. Schmitt, and H. van Hasselt, “Multi-task deep reinforcement learning with popart,” 2018
2018
Cited alongside, same era.
J. S. Matthis, J. L. Yates, and M. M. Hayhoe, “Gaze and the control of foot placement when walking in natural terrain,” Current Biology , vol. 28, no. 8, pp. 1224 – 1233.e5, 2018. [Online]. Available: http://www.sciencedirect.com/science/article/pii/S0960982218303099
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, 2019
2019
Later among the works it cites.
D. Chen, B. Zhou, V. Koltun, and P. Krähenbühl, “Learning by cheating,” 2019
2019
Later among the works it cites.
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine, “Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning,” 2019
2019
Later among the works it cites.
T. Yu, S. Jumar, A. Gupta, S. Levine, K. Hausmann, and C. Finn, “Multi-task reinforcement learning without interference,” 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Unitree, “Laikago: Let’s challenge new possibilities,” 2018. [Online]. Available: http://www.unitree.cc/e/action/ShowInfo.php?classid=6&id=1
2018
Cited alongside, same era.
J. Schulman, P. Moritz, S. Levine, M. Jordan, and P. Abbeel, “High-dimensional continuous control using generalized advantage estimation,” 2018
2018
Cited alongside, same era.
E. Coumans and Y. Bai, “Pybullet, a python module for physics simulation for games, robotics and machine learning,” http://pybullet.org, 2016–2019
2019
Cited alongside, same era.
R. Grandia, F. Farshidian, R. Ranftl, and M. Hutter, “Feedback mpc for torque-controlled legged robots,” 05 2019
2019
Cited alongside, same era.
T. Haarnoja, S. Ha, A. Zhou, J. Tan, G. Tucker, and S. Levine, “Learning to walk via deep reinforcement learning,” in Robotics: Science and Systems , 2019
2019
Cited alongside, same era.
2019
Later among the works it cites.
K. Albee, A. C. Hernandez, O. Jia-Richards, and A. T. Espinoza, “Real-time motion planning in unknown environments for legged robotic planetary exploration,” in 2020 IEEE Aerospace Conference , 2020, pp. 1–9
2020
Closest in time.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science Robotics , vol. 5, no. 47, 2020. [Online]. Available: https://robotics.sciencemag.org/content/5/47/eabc5986
2020
Closest in time.
S. Ha, P. Xu, Z. Tan, S. Levine, and J. Tan, “Learning to walk in the real world with minimal human effort,” 2020
2020
Closest in time.
V. Tsounis, M. Alge, J. Lee, F. Farshidian, and M. Hutter, “Deepgait: Planning and control of quadrupedal gaits using deep reinforcement learning,” 2020
2020
Closest in time.
M. Andrychowicz, A. Raichuk, P. Stańczyk, M. Orsini, S. Girgin, R. Marinier, L. Hussenot, M. Geist, O. Pietquin, M. Michalski, S. Gelly, and O. Bachem, “What matters in on-policy reinforcement learning? a large-scale empirical study,” 2020
2020
Closest in time.