Fetching the paper…
Reading the bibliography…
The ability to walk in new scenarios is a key milestone on the path toward real-world applications of legged robots.
1903
Earlier work this paper cites.
1903
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International Conference on Machine Learning , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
N. Hansen, A. Ostermeier, and A. Gawelczyk, “On the adaptation of arbitrary normal mutation distributions in evolution strategies: The generating set adaptation.” in ICGA , 1995, pp. 57–64
1995
Earlier work this paper cites.
J. Mockus, Bayesian approach to global optimization: theory and applications . Springer Science & Business Media, 2012, vol. 37
2012
Earlier work this paper cites.
A. Aswani, P. Bouffard, and C. Tomlin, “Extensions of learning-based model predictive control for real-time application to a quadrotor helicopter,” in 2012 American Control Conference (ACC) . IEEE, 2012, pp. 4661–4666
2012
Earlier work this paper cites.
M. Tanaskovic, L. Fagiano, R. Smith, P. Goulart, and M. Morari, “Adaptive model predictive control for constrained linear systems,” in 2013 European Control Conference (ECC) . IEEE, 2013, pp. 382–387
2013
Earlier work this paper cites.
P. Manganiello, M. Ricco, G. Petrone, E. Monmasson, and G. Spagnuolo, “Optimization of perturbative pv mppt methods through online system identification,” IEEE Transactions on Industrial Electronics , vol. 61, no. 12, pp. 6812–6821, 2014
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. Cully, J. Clune, D. Tarapore, and J.-B. Mouret, “Robots that can adapt like animals,” Nature , vol. 521, no. 7553, p. 503, 2015
2015
Earlier work this paper cites.
I. Lenz, R. A. Knepper, and A. Saxena, “Deepmpc: Learning deep latent features for model predictive control.” in Robotics: Science and Systems . Rome, Italy, 2015
2015
Earlier work this paper cites.
S. J. Wright, “Coordinate descent algorithms,” Mathematical Programming , vol. 151, no. 1, pp. 3–34, 2015
2015
Earlier work this paper cites.
G. Kenneally, A. De, and D. E. Koditschek, “Design principles for a family of direct-drive legged robots,” IEEE Robotics and Automation Letters , vol. 1, no. 2, pp. 900–907, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in Intelligent Robots and Systems (IROS), 2017 IEEE/RSJ International Conference on . IEEE, 2017, pp. 23–30
2017
Cited alongside, same era.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” in Proceedings of the 34th International Conference on Machine Learning-Volume 70 . JMLR. org, 2017, pp. 1126–1135
2017
Cited alongside, same era.
M. Neunert, T. Boaventura, and J. Buchli, “Why off-the-shelf physics simulators fail in evaluating feedback controller performance-a case study for quadrupedal robots,” in Advances in Cooperative Robotics . World Scientific, 2017, pp. 464–472
2017
Cited alongside, same era.
J. P. Hanna and P. Stone, “Grounded action transformation for robot learning in simulation,” in Thirty-First AAAI Conference on Artificial Intelligence , 2017
2017
K. Fang, Y. Bai, S. Hinterstoisser, S. Savarese, and M. Kalakrishnan, “Multi-task domain adaptation for deep learning of instance grasping from simulation,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 3516–3523
2018
Later among the works it cites.
R. Houthooft, Y. Chen, P. Isola, B. Stadie, F. Wolski, O. J. Ho, and P. Abbeel, “Evolved policy gradients,” in Advances in Neural Information Processing Systems , 2018, pp. 5400–5409
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning.” ICML, 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
W. Yu, J. Tan, C. K. Liu, and G. Turk, “Preparing for the unknown: Learning a universal policy with online system identification,” in Proceedings of Robotics: Science and Systems , Cambridge, Massachusetts, July 2017
2017
Cited alongside, same era.
E. Coumans and Y. Bai, “Pybullet, a python module for physics simulation in robotics, games and machine learning.” 2016-2017. [Online]. Available: http://pybullet.org
2017
Cited alongside, same era.
P.-L. Bacon, J. Harb, and D. Precup, “The option-critic architecture,” in Thirty-First AAAI Conference on Artificial Intelligence , 2017
2017
Cited alongside, same era.
L. Liu and J. Hodgins, “Learning to schedule control fragments for physics-based characters using deep q-learning,” ACM Transactions on Graphics (TOG) , vol. 36, no. 3, p. 29, 2017
2017
Cited alongside, same era.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” in Proceedings of Robotics: Science and Systems , Pittsburgh, Pennsylvania, June 2018
2018
Cited alongside, same era.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . ICRA, 2018, pp. 1–8
2018
Cited alongside, same era.
A. Rai, R. Antonova, S. Song, W. Martin, H. Geyer, and C. Atkeson, “Bayesian optimization using domain knowledge on the atrias biped,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 1771–1778
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
J. Lee, M. Grey, S. Ha, T. Kunz, S. Jain, Y. Ye, S. Srinivasa, M. Stilman, and C. Liu, “Dart: Dynamic animation and robotics toolkit,” The Journal of Open Source Software , vol. 3, p. 500, 02 2018
2018
Later among the works it cites.
2019
Closest in time.
W. Yu, C. K. Liu, and G. Turk, “Policy transfer with strategy optimization,” in International Conference on Learning Representations , 2019. [Online]. Available: https://openreview.net/forum?id=H1g6osRcFQ
2019
Closest in time.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, p. eaau5872, 2019
2019
Closest in time.
2019
Closest in time.
S. James, P. Wohlhart, M. Kalakrishnan, D. Kalashnikov, A. Irpan, J. Ibarz, S. Levine, R. Hadsell, and K. Bousmalis, “Sim-to-real via sim-to-sim: Data-efficient robotic grasping via randomized-to-canonical adaptation networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 12 627–12 637
2019
Closest in time.
K. Varelas, “Benchmarking large scale variants of cma-es and l-bfgs-b on the bbob-largescale testbed,” 2019
2019
Closest in time.