Fetching the paper…
Reading the bibliography…
Learning to locomote to arbitrary goals on hardware remains a challenging problem for reinforcement learning.
1905
Earlier work this paper cites.
R. S. Sutton, D. Precup, and S. Singh, “Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,” Artificial intelligence , vol. 112, no. 1-2, pp. 181–211, 1999
1999
Earlier work this paper cites.
D. Q. Mayne, J. B. Rawlings, C. V. Rao, and P. O. Scokaert, “Constrained model predictive control: Stability and optimality,” Automatica , vol. 36, no. 6, pp. 789–814, 2000
2000
Earlier work this paper cites.
H. Kimura, Y. Fukuoka, and A. H. Cohen, “Biologically inspired adaptive walking of a quadruped robot,” Philosophical Transactions of the Royal Society A: Mathematical, Physical and Engineering Sciences , vol. 365, no. 1850, pp. 153–170, 2006
2006
Earlier work this paper cites.
A. Crespi, D. Lachat, A. Pasquier, and A. J. Ijspeert, “Controlling swimming and crawling in a fish robot using a central pattern generator,” Autonomous Robots , vol. 25, no. 1-2, pp. 3–13, 2008
2008
Earlier work this paper cites.
A. Herdt, H. Diedam, P.-B. Wieber, D. Dimitrov, K. Mombaur, and M. Diehl, “Online walking motion generation with automatic footstep placement,” Advanced Robotics , vol. 24, no. 5-6, pp. 719–737, 2010
2010
Earlier work this paper cites.
F. Stulp and S. Schaal, “Hierarchical reinforcement learning with movement primitives,” in 2011 11th IEEE-RAS International Conference on Humanoid Robots . IEEE, 2011, pp. 231–238
2011
Earlier work this paper cites.
C. Daniel, G. Neumann, O. Kroemer, and J. Peters, “Learning sequential motor tasks,” in 2013 IEEE International Conference on Robotics and Automation . IEEE, 2013, pp. 2626–2632
2013
Earlier work this paper cites.
Y. Fukuoka, Y. Habu, and T. Fukui, “Analysis of the gait generation principle by a simulated quadruped model with a cpg incorporating vestibular modulation,” Biological cybernetics , vol. 107, no. 6, pp. 695–710, 2013
2013
Earlier work this paper cites.
H.-W. Park, P. M. Wensing, S. Kim et al. , “Online planning for autonomous running jumps over obstacles in high-speed quadrupeds,” 2015
2015
Earlier work this paper cites.
J. Koenemann, A. Del Prete, Y. Tassa, E. Todorov, O. Stasse, M. Bennewitz, and N. Mansard, “Whole-body model-predictive control applied to the hrp-2 humanoid,” in 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2015, pp. 3346–3351
2015
Earlier work this paper cites.
J. André, C. Teixeira, C. P. Santos, and L. Costa, “Adapting biped locomotion to sloped environments,” Journal of Intelligent & Robotic Systems , vol. 80, no. 3-4, pp. 625–640, 2015
2015
Earlier work this paper cites.
A. Cully, J. Clune, D. Tarapore, and J.-B. Mouret, “Robots that can adapt like animals,” Nature , vol. 521, no. 7553, pp. 503–507, 2015
2015
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International conference on machine learning , 2015, pp. 1889–1897
2015
Cited alongside, same era.
S. Feng, “Online hierarchical optimization for humanoid control,” 2016
2016
Cited alongside, same era.
C. Daniel, H. Van Hoof, J. Peters, and G. Neumann, “Probabilistic inference for determining options in reinforcement learning,” Machine Learning , vol. 104, no. 2-3, pp. 337–357, 2016
2016
Cited alongside, same era.
P.-L. Bacon, J. Harb, and D. Precup, “The option-critic architecture,” 2016
2016
Cited alongside, same era.
X. B. Peng, G. Berseth, and M. van de Panne, “Terrain-adaptive locomotion skills using deep reinforcement learning,” ACM Trans. Graph. , vol. 35, no. 4, pp. 81:1–81:12, Jul. 2016. [Online]. Available: http://doi.acm.org/10.1145/2897824.2925881
2018
Later among the works it cites.
S. Mason, N. Rotella, S. Schaal, and L. Righetti, “An mpc walking framework with external contact forces,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 1785–1790
2018
Later among the works it cites.
B. Yang, G. Wang, R. Calandra, D. Contreras, S. Levine, and K. Pister, “Learning flexible and reusable locomotion primitives for a microrobot,” IEEE Robotics and Automation Letters (RA-L) , vol. 3, no. 3, pp. 1904–1911, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
X. B. Peng, G. Berseth, K. Yin, and M. Van De Panne, “Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning,” ACM Transactions on Graphics (TOG) , vol. 36, no. 4, p. 41, 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
M. Andrychowicz, F. Wolski, A. Ray, J. Schneider, R. Fong, P. Welinder, B. McGrew, J. Tobin, O. P. Abbeel, and W. Zaremba, “Hindsight experience replay,” in Advances in Neural Information Processing Systems , 2017, pp. 5048–5058
2017
Cited alongside, same era.
D. Owaki and A. Ishiguro, “A quadruped robot exhibiting spontaneous gait transitions from walking to trotting to galloping,” Scientific reports , vol. 7, no. 1, p. 277, 2017
2017
Cited alongside, same era.
X. B. Peng, G. Berseth, K. Yin, and M. van de Panne, “Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning,” ACM Transactions on Graphics (Proc. SIGGRAPH 2017) , vol. 36, no. 4, 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
K. Chua, R. Calandra, R. McAllister, and S. Levine, “Deep reinforcement learning in a handful of trials using probabilistic dynamics models,” in Advances in Neural Information Processing Systems , 2018, pp. 4754–4765
2018
Cited alongside, same era.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Later among the works it cites.
J. Di Carlo, P. M. Wensing, B. Katz, G. Bledt, and S. Kim, “Dynamic locomotion in the mit cheetah 3 through convex model-predictive control,” in 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2018, pp. 1–9
2018
Later among the works it cites.
T. Li, H. Geyer, C. G. Atkeson, and A. Rai, “Using deep reinforcement learning to learn high-level policies on the atrias biped,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 263–269
2019
Closest in time.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, p. eaau5872, 2019
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
“Pybullet simulator,” https://github.com/bulletphysics/bullet3, accessed: 2019-09
2019
Closest in time.