Fetching the paper…
Reading the bibliography…
We address the problem of enabling quadrupedal robots to perform precise shooting skills in the real world using reinforcement learning.
M. Veloso, W. Uther, M. Fijita, M. Asada, and H. Kitano, “Playing soccer with legged robots,” in Proc. Int. Conf. Intell. Robots Syst. , 1998
1998
Earlier work this paper cites.
P. Stone, “Intelligent autonomous robotics: A robot soccer case study,” Synth. Lect. Artif. Intell. Machine Learn. , 2007
2007
Earlier work this paper cites.
S. K. Chalup, C. L. Murch, and M. J. Quinlan, “Machine learning with aibo robots in the four-legged league of robocup,” Trans. Syst., Man, and Cybern., Part C , 2007
2007
Earlier work this paper cites.
J. Peters and S. Schaal, “Reinforcement learning of motor skills with policy gradients,” Neural Networks , 2008
2008
Earlier work this paper cites.
M. Friedmann, J. Kiener, S. Petters, D. Thomas, O. Von Stryk, and H. Sakamoto, “Versatile, high-quality motions and behavior control of a humanoid soccer robot,” Int. J. Human. Robot. , 2008
2008
Earlier work this paper cites.
S. Behnke and J. Stückler, “Hierarchical reactive control for humanoid soccer robots,” Int. J. Human. Robot. , 2008
2008
Earlier work this paper cites.
C. A. Acosta-Calderon, R. E. Mohan, C. Zhou, L. Hu, P. K. Yue, and H. Hu, “A modular architecture for humanoid soccer robots with distributed behavior control,” Int. J. Human. Robot. , 2008
2008
Earlier work this paper cites.
J.-w. Choi, R. Curry, and G. Elkaim, “Path planning based on bézier curve for autonomous ground vehicles,” in World Congress on Engineering and Computer Science , 2008
2008
Earlier work this paper cites.
A. Cherubini, F. Giannone, L. Iocchi, D. Nardi, and P. F. Palamara, “Policy gradient learning for quadruped soccer robots,” Robot. Auton. Syst. , 2010
2010
Earlier work this paper cites.
J. Kober, E. Oztop, and J. Peters, “Reinforcement learning to adjust robot movements to new situations,” in Twenty-Second International Joint Conference on Artificial Intelligence , 2011
2011
Earlier work this paper cites.
E. Olson, “Apriltag: A robust and flexible visual fiducial system,” in Proc. Int. Conf. Robot. Automat. , 2011
2011
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in Proc. Int. Conf. Intell. Robots Syst. , 2012
2012
Earlier work this paper cites.
J. Peters, J. Kober, K. Mülling, O. Krämer, and G. Neumann, “Towards robot skill learning: From simple skills to table tennis,” in Machine Learning and Knowledge Discovery in Databases , 2013
2013
Cited alongside, same era.
M. P. Deisenroth, P. Englert, J. Peters, and D. Fox, “Multi-task policy search for robotics,” in Proc. Int. Conf. Robot. Automat. , 2014
2014
Cited alongside, same era.
N. Jouandeau and V. Hugel, “Optimization of parametrised kicking motion for humanoid soccer player,” in Proc. Int. Conf. Auton. Robot Syst. Compet. , 2014
2014
Cited alongside, same era.
A. Ghadirzadeh, A. Maki, D. Kragic, and M. Björkman, “Deep predictive policy training using reinforcement learning,” in Proc. Int. Conf. Intell. Robots Syst. , 2017
2017
Cited alongside, same era.
Y. Chebotar, K. Hausman, M. Zhang, G. Sukhatme, S. Schaal, and S. Levine, “Combining model-based and model-free updates for trajectory-centric reinforcement learning,” in Proc. Int. Conf. Machine Learn. , 2017
X. B. Peng, E. Coumans, T. Zhang, T.-W. Lee, J. Tan, and S. Levine, “Learning agile robotic locomotion skills by imitating animals,” Robotics: Science and Systems , 2020
2020
Later among the works it cites.
H. Teixeira, T. Silva, M. Abreu, and L. P. Reis, “Humanoid robot kick in motion ability for playing robotic soccer,” in Proc. Int. Conf. Auton. Robot Syst. Compet. , 2020
2020
Later among the works it cites.
S. Gilroy, D. Lau, L. Yang, E. Izaguirre, K. Biermayer, A. Xiao, M. Sun, A. Agrawal, J. Zeng, Z. Li et al. , “Autonomous navigation for quadrupedal robots with optimized jumping through constrained obstacles,” in Proc. Int. Conf. Automat. Sci. Eng. , 2021
2021
Later among the works it cites.
F. Shi, T. Homberger, J. Lee, T. Miki, M. Zhao, F. Farshidian, K. Okada, M. Inaba, and M. Hutter, “Circus anymal: A quadruped learning dexterous manipulation with its limbs,” in Proc. Int. Conf. Robot. Automat. , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
X. B. Peng, G. Berseth, K. Yin, and M. Van De Panne, “Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning,” ACM Transactions on Graphics , 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
A. W. Winkler, C. D. Bellicoso, M. Hutter, and J. Buchli, “Gait and trajectory optimization for legged systems through phase-based end-effector parameterization,” Robot. Automat. Lett. , 2018
2018
Cited alongside, same era.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in Proc. Int. Conf. Robot. Automat. , 2018
2018
Cited alongside, same era.
2019
Cited alongside, same era.
A. Zeng, S. Song, J. Lee, A. Rodriguez, and T. Funkhouser, “Tossingbot: Learning to throw arbitrary objects with residual physics,” Robotics: Science and Systems , 2019
2019
Cited alongside, same era.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , 2019
2019
Cited alongside, same era.
A. Kumar, Z. Fu, D. Pathak, and J. Malik, “Rma: Rapid motor adaptation for legged robots,” Robotics: Science and Systems , 2021
2021
Later among the works it cites.
Z. Li, X. Cheng, X. B. Peng, P. Abbeel, S. Levine, G. Berseth, and K. Sreenath, “Reinforcement learning for robust parameterized locomotion control of bipedal robots,” in Proc. Int. Conf. Robot. Automat. , 2021
2021
Later among the works it cites.
I. J. da Silva, D. H. Perico, T. P. D. Homem, and R. A. da Costa Bianchi, “Deep reinforcement learning for a humanoid robot soccer player,” J. Intell. Robot. Syst. , 2021
2021
Later among the works it cites.
X. Chen, C. Wang, Z. Zhou, and K. W. Ross, “Randomized ensembled double q-learning: Learning fast without a model,” in Proc. Int. Conf. Learn. Repres. , 2021
2021
Later among the works it cites.
Y. Yang, T. Zhang, E. Coumans, J. Tan, and B. Boots, “Fast and efficient locomotion via learned gait transitions,” in Proc. Conf. Robot Learn. , 2021
2021
Later among the works it cites.
F. Muratore, T. Gruner, F. Wiese, B. Belousov, M. Gienger, and J. Peters, “Neural posterior domain randomization,” in Proc. Conf. Robot Learn. , 2022
2022
Closest in time.
L. Smith, J. C. Kew, X. B. Peng, S. Ha, J. Tan, and S. Levine, “Legged robots that keep on learning: Fine-tuning locomotion policies in the real world,” in Proc. Int. Conf. Robot. Automat. , 2022
2022
Closest in time.
G. Ji, J. Mun, H. Kim, and J. Hwangbo, “Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion,” Robot. Automat. Lett. , 2022
2022
Closest in time.