Fetching the paper…
Reading the bibliography…
On-robot Reinforcement Learning is a promising approach to train embodiment-aware policies for legged robots.
I. Selesnick and C. Burrus, “Generalized digital butterworth filter design,” Transactions on signal processing , 1998
1998
Earlier work this paper cites.
E. Marder and D. Bucher, “Central pattern generators and the control of rhythmic movements,” Current biology , vol. 11, no. 23, pp. R986–R996, 2001
2001
Earlier work this paper cites.
J. Pratt, J. Carff, S. Drakunov, and A. Goswami, “Capture point: A step toward humanoid push recovery,” in International conference on humanoid robots , 2006
2006
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, A. K. Dey et al. , “Maximum entropy inverse reinforcement learning.” in Aaai , 2008
2008
Earlier work this paper cites.
A. J. Ijspeert, “Central pattern generators for locomotion control in animals and robots: a review,” Neural networks , vol. 21, no. 4, pp. 642–653, 2008
2008
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, and J. Weston, “Curriculum learning,” in International conference on machine learning , 2009
2009
Earlier work this paper cites.
G. Casiez, N. Roussel, and D. Vogel, “1 € filter: a simple speed-based low-pass filter for noisy input in interactive systems,” in Conference on human factors in computing systems , 2012
2012
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in International conference on intelligent robots and systems , 2012
2012
Earlier work this paper cites.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in International conference on machine learning , 2015
2015
Earlier work this paper cites.
S. Kuindersma, R. Deits, M. Fallon, A. Valenzuela, H. Dai, F. Permenter, T. Koolen, P. Marion, and R. Tedrake, “Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,” Autonomous robots , vol. 40, pp. 429–455, 2016
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in International conference on intelligent robots and systems , 2017
2017
Earlier work this paper cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International conference on machine learning , 2018
2018
Earlier work this paper cites.
J. S. Furtado, H. H. Liu, G. Lai, H. Lacheray, and J. Desouza-Coelho, “Comparative analysis of optitrack motion capture systems,” in Advances in Motion Sensing and Control for Robotic Applications . Springer, 2019, pp. 15–31
2019
Earlier work this paper cites.
X. Chen, C. Wang, Z. Zhou, and K. Ross, “Randomized ensembled double Q-learning: Learning fast without a model,” in International conference on learning representations , 2021
2021
Cited alongside, same era.
T. Hiraoka, T. Imagawa, T. Hashimoto, T. Onishi, and Y. Tsuruoka, “Dropout q-functions for doubly efficient reinforcement learning,” in International conference on learning representations , 2021
2021
Cited alongside, same era.
Y. Shao, Y. Jin, X. Liu, W. He, H. Wang, and W. Yang, “Learning free gait transition for quadruped robots via phase-guided controller,” Robotics and automation letters , vol. 7, no. 2, pp. 1230–1237, 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
A. Agarwal, A. Kumar, J. Malik, and D. Pathak, “Legged locomotion in challenging terrains using egocentric vision,” in Conference on robot learning , 2023
2023
Later among the works it cites.
L. Smith, I. Kostrikov, and S. Levine, “A walk in the park: Learning to walk in 20 minutes with model-free reinforcement learning,” in Robotics: Science and systems , 2023
2023
Later among the works it cites.
N. Bohlinger and K. Dorer, “Rl-x: A deep reinforcement learning library (not only) for robocup,” in Robot world cup . Springer, 2023, pp. 228–239
2023
Later among the works it cites.
N. Bohlinger, G. Czechmanowski, M. Krupka, P. Kicki, K. Walas, J. Peters, and D. Tateo, “One policy to run them all: an end-to-end learning approach to multi-embodiment locomotion,” in Conference on robot learning , 2024
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
G. Bellegarda, Y. Chen, Z. Liu, and Q. Nguyen, “Robust high-speed running for quadruped robots via deep reinforcement learning,” in International conference on intelligent robots and systems , 2022
2022
Cited alongside, same era.
N. Rudin, D. Hoeller, P. Reist, and M. Hutter, “Learning to walk in minutes using massively parallel deep reinforcement learning,” in Conference on robot learning , 2022
2022
Cited alongside, same era.
G. Ji, J. Mun, H. Kim, and J. Hwangbo, “Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion,” Robotics and automation letters , vol. 7, no. 2, pp. 4630–4637, 2022
2022
Cited alongside, same era.
E. Nikishin, M. Schwarzer, P. D’Oro, P.-L. Bacon, and A. Courville, “The primacy bias in deep reinforcement learning,” in International conference on machine learning , 2022
2022
Cited alongside, same era.
P. D’Oro, M. Schwarzer, E. Nikishin, P.-L. Bacon, M. G. Bellemare, and A. Courville, “Sample-efficient reinforcement learning by breaking the replay ratio barrier,” in International conference on learning representations , 2022
2022
Cited alongside, same era.
N. Rudin, D. Hoeller, P. Reist, and M. Hutter, “Learning to walk in minutes using massively parallel deep reinforcement learning,” in Conference on robot learning , 2022
2022
Cited alongside, same era.
Y. Wu, X. Chen, C. Wang, Y. Zhang, and K. W. Ross, “Aggressive q-learning with ensembles: Achieving both high sample efficiency and high asymptotic performance,” in Deep reinforcement learning workshop @ NeurIPS , 2022
2022
Cited alongside, same era.
Z. Zhuang, Z. Fu, J. Wang, C. Atkeson, S. Schwertfeger, C. Finn, and H. Zhao, “Robot parkour learning,” in Conference on robot learning , 2023
2023
Cited alongside, same era.
G. B. Margolis, G. Yang, K. Paigwar, T. Chen, and P. Agrawal, “Rapid locomotion via reinforcement learning,” International journal of robotics research , vol. 43, no. 4, pp. 572–587, 2024
2024
Later among the works it cites.
C. Zhang, N. Rudin, D. Hoeller, and M. Hutter, “Learning agile locomotion on risky terrains,” in International conference on intelligent robots and systems , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
L. Smith, Y. Cao, and S. Levine, “Grow your limits: Continuous improvement with real-world rl for robotic locomotion,” in International conference on robotics and automation , 2024, pp. 10 829–10 836
2024
Later among the works it cites.
J. Levy, T. Westenbroek, and D. Fridovich-Keil, “Learning to walk from three minutes of real-world data with semi-structured dynamics models,” in Conference on robot learning , 2024
2024
Later among the works it cites.
A. Bhatt, D. Palenicek, B. Belousov, M. Argus, A. Amiranashvili, T. Brox, and J. Peters, “CrossQ: Batch normalization in deep reinforcement learning for greater sample efficiency and simplicity,” in International conference on learning representations , 2024
2024
Later among the works it cites.
2025
Closest in time.
2025
Closest in time.