Fetching the paper…
Reading the bibliography…
Deep Reinforcement Learning (RL) has emerged as a promising method to develop humanoid robot locomotion controllers.
C. E. Garcia, D. M. Prett, and M. Morari, “Model predictive control: Theory and practice—a survey,” Automatica , vol. 25, no. 3, pp. 335–348, 1989
1989
Earlier work this paper cites.
A. Müller, “Integral probability metrics and their generating classes of functions,” Advances in Applied Probability , vol. 29, no. 2, pp. 429–443, 1997
1997
Earlier work this paper cites.
Y. Sun, D. Wierstra, T. Schaul, and J. Schmidhuber, “Efficient natural evolution strategies,” in Proceedings of the 11th Annual conference on Genetic and evolutionary computation , 2009, pp. 539–546
2009
Earlier work this paper cites.
Y. Zheng, M. C. Lin, D. Manocha, A. H. Adiwahono, and C.-M. Chew, “A walking pattern generator for biped robots on uneven terrains,” in 2010 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2010, pp. 4483–4488
2010
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ international conference on intelligent robots and systems . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
N. Van der Noot and A. Barrea, “Zero-moment point on a bipedal robot under bio-inspired walking control,” in MELECON 2014-2014 17th IEEE Mediterranean Electrotechnical Conference . IEEE, 2014, pp. 85–90
2014
Earlier work this paper cites.
S. Caron, Q.-C. Pham, and Y. Nakamura, “Stability of surface contacts for humanoid robots: Closed-form formulae of the contact wrench cone for rectangular support areas,” in 2015 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2015, pp. 5107–5112
2015
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International conference on machine learning . PMLR, 2015, pp. 1889–1897
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
H. Park, B.-Y. Lee, M.-J. Tahk, and D.-W. Yoo, “Differential game based air combat maneuver generation using scoring function matrix,” International Journal of Aeronautical and Space Sciences , vol. 17, no. 2, pp. 204–213, 2016
2016
Earlier work this paper cites.
F. Farshidian, E. Jelavic, A. Satapathy, M. Giftthaler, and J. Buchli, “Real-time motion planning of legged robots: A model predictive control approach,” in 2017 IEEE-RAS 17th International Conference on Humanoid Robotics (Humanoids) . IEEE, 2017, pp. 577–584
2017
Earlier work this paper cites.
M. G. Bellemare, W. Dabney, and R. Munos, “A distributional perspective on reinforcement learning,” in International Conference on Machine Learning . PMLR, 2017, pp. 449–458
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
X. B. Peng and M. Van De Panne, “Learning locomotion skills using deeprl: Does the choice of action space matter?” in Proceedings of the ACM SIGGRAPH/Eurographics Symposium on Computer Animation , 2017, pp. 1–13
2017
Earlier work this paper cites.
X. B. Peng, P. Abbeel, S. Levine, and M. Van de Panne, “Deepmimic: Example-guided deep reinforcement learning of physics-based character skills,” ACM Transactions On Graphics (TOG) , vol. 37, no. 4, pp. 1–14, 2018
2018
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Cited alongside, same era.
W. Dabney, G. Ostrovski, D. Silver, and R. Munos, “Implicit quantile networks for distributional reinforcement learning,” in International conference on machine learning . PMLR, 2018, pp. 1096–1105
2018
Cited alongside, same era.
W. Dabney, M. Rowland, M. Bellemare, and R. Munos, “Distributional reinforcement learning with quantile regression,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 32, no. 1, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
N. Rudin, D. Hoeller, P. Reist, and M. Hutter, “Learning to walk in minutes using massively parallel deep reinforcement learning,” in Conference on Robot Learning . PMLR, 2022, pp. 91–100
2022
Later among the works it cites.
Y. Fuchioka, Z. Xie, and M. Van de Panne, “Opt-mimic: Imitation of optimized trajectories for dynamic quadruped behaviors,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 5092–5098
2023
Later among the works it cites.
F. Semeraro, A. Griffiths, and A. Cangelosi, “Human–robot collaboration and machine learning: A systematic review of recent research,” Robotics and Computer-Integrated Manufacturing , vol. 79, p. 102432, 2023
2023
Later among the works it cites.
Z. Luo, J. Cao, K. Kitani, W. Xu et al. , “Perpetual humanoid control for real-time simulated avatars,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 10 895–10 904
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International conference on machine learning . PMLR, 2018, pp. 1861–1870
2018
Cited alongside, same era.
W. Yu, G. Turk, and C. K. Liu, “Learning symmetric and low-energy locomotion,” ACM Transactions on Graphics (TOG) , vol. 37, no. 4, pp. 1–12, 2018
2018
Cited alongside, same era.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, p. eaau5872, 2019
2019
Cited alongside, same era.
K. Bergamin, S. Clavet, D. Holden, and J. R. Forbes, “Drecon: data-driven responsive control of physics-based characters,” ACM Transactions On Graphics (TOG) , vol. 38, no. 6, pp. 1–11, 2019
2019
Cited alongside, same era.
O. M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray et al. , “Learning dexterous in-hand manipulation,” The International Journal of Robotics Research , vol. 39, no. 1, pp. 3–20, 2020
2020
Cited alongside, same era.
E. Valassakis, Z. Ding, and E. Johns, “Crossing the gap: A deep dive into zero-shot sim-to-real transfer for dynamics,” in 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2020, pp. 5372–5379
2020
Cited alongside, same era.
F. Zhou, J. Wang, and X. Feng, “Non-crossing quantile regression for distributional reinforcement learning,” Advances in Neural Information Processing Systems , vol. 33, pp. 15 909–15 919, 2020
2020
Cited alongside, same era.
D. W. Nam, Y. Kim, and C. Y. Park, “Gmac: A distributional perspective on actor-critic framework,” in International Conference on Machine Learning . PMLR, 2021, pp. 7927–7936
2021
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Tong, H. Liu, and Z. Zhang, “Advancements in humanoid robots: A comprehensive review and future prospects,” IEEE/CAA Journal of Automatica Sinica , vol. 11, no. 2, pp. 301–328, 2024
2024
Later among the works it cites.
D. Vernon and G. Sandini, “The importance of being humanoid,” International Journal of Humanoid Robotics , vol. 21, no. 01, 2024
2024
Later among the works it cites.
I. Radosavovic, T. Xiao, B. Zhang, T. Darrell, J. Malik, and K. Sreenath, “Real-world humanoid locomotion with reinforcement learning,” Science Robotics , vol. 9, no. 89, p. eadi9579, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
Q. Zhang, P. Cui, D. Yan, J. Sun, Y. Duan, G. Han, W. Zhao, W. Zhang, Y. Guo, A. Zhang et al. , “Whole-body humanoid robot locomotion with human reference,” in 2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2024, pp. 11 225–11 231
2024
Later among the works it cites.
2024
Later among the works it cites.
L. Campanaro, S. Gangapurwala, W. Merkt, and I. Havoutis, “Learning and deploying robust locomotion policies with minimal dynamics randomization,” in 6th Annual Learning for Dynamics & Control Conference . PMLR, 2024, pp. 578–590
2024
Later among the works it cites.
Z. Li, X. B. Peng, P. Abbeel, S. Levine, G. Berseth, and K. Sreenath, “Reinforcement learning for versatile, dynamic, and robust bipedal locomotion control,” The International Journal of Robotics Research , p. 02783649241285161, 2024
2024
Later among the works it cites.