Fetching the paper…
Reading the bibliography…
The robustness of legged locomotion is crucial for quadrupedal robots in challenging terrains.
G. Feng, H. Zhang, Z. Li, X. B. Peng, B. Basireddy, L. Yue, Z. Song, L. Yang, Y. Liu, K. Sreenath, et al. , “Genloco: Generalized locomotion controllers for quadrupedal robots,” in Conference on Robot Learning (CoRL) . PMLR, 2023, pp. 1893–1903
1903
Earlier work this paper cites.
S. C. Jaquette, “Markov decision processes with a new optimality criterion: Discrete time,” The Annals of Statistics , vol. 1, no. 3, pp. 496–505, 1973
1973
Earlier work this paper cites.
M. J. Sobel, “The variance of discounted markov decision processes,” Journal of Applied Probability , vol. 19, no. 4, pp. 794–802, 1982
1982
Earlier work this paper cites.
D. J. White, “Mean, variance, and probabilistic criteria in finite markov decision processes: A review,” Journal of Optimization Theory and Applications , vol. 56, pp. 1–29, 1988
1988
Earlier work this paper cites.
A. Müller, “Integral probability metrics and their generating classes of functions,” Advances in applied probability , vol. 29, no. 2, pp. 429–443, 1997
1997
Earlier work this paper cites.
R. T. Rockafellar, S. Uryasev, et al. , “Optimization of conditional value-at-risk,” Journal of risk , vol. 2, pp. 21–42, 2000
2000
Earlier work this paper cites.
Y. Shen, M. J. Tobia, T. Sommer, and K. Obermayer, “Risk-sensitive reinforcement learning,” Neural computation , vol. 26, no. 7, pp. 1298–1328, 2014
2014
Earlier work this paper cites.
Y. Chow and M. Ghavamzadeh, “Algorithms for cvar optimization in mdps,” Advances in neural information processing systems , vol. 27, 2014
2014
Earlier work this paper cites.
L. Prashanth, C. Jie, M. Fu, S. Marcus, and C. Szepesvári, “Cumulative prospect theory meets reinforcement learning: Prediction and control,” in International Conference on Machine Learning . PMLR, 2016, pp. 1406–1415
2016
Earlier work this paper cites.
J. Schulman, P. Moritz, S. Levine, M. I. Jordan, and P. Abbeel, “High-dimensional continuous control using generalized advantage estimation,” in 4th International Conference on Learning Representations (ICLR) , 2016
2016
Earlier work this paper cites.
M. G. Bellemare, W. Dabney, and R. Munos, “A distributional perspective on reinforcement learning,” in International conference on Machine Learning . PMLR, 2017, pp. 449–458
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. W. Winkler, C. D. Bellicoso, M. Hutter, and J. Buchli, “Gait and trajectory optimization for legged systems through phase-based end-effector parameterization,” IEEE Robotics and Automation Letters , vol. 3, no. 3, pp. 1560–1567, 2018
2018
Earlier work this paper cites.
S. Depeweg, J.-M. Hernandez-Lobato, F. Doshi-Velez, and S. Udluft, “Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning,” in International Conference on Machine Learning . PMLR, 2018, pp. 1184–1193
2018
Earlier work this paper cites.
W. Dabney, M. Rowland, M. Bellemare, and R. Munos, “Distributional reinforcement learning with quantile regression,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 32, no. 1, 2018
2018
Earlier work this paper cites.
W. Dabney, G. Ostrovski, D. Silver, and R. Munos, “Implicit quantile networks for distributional reinforcement learning,” in International conference on Machine Learning . PMLR, 2018, pp. 1096–1105
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, p. eaau5872, 2019
2019
Earlier work this paper cites.
T. Haarnoja, S. Ha, A. Zhou, J. Tan, G. Tucker, and S. Levine, “Learning to walk via deep reinforcement learning,” in Robotics: Science and Systems , 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
B. Mavrin, H. Yao, L. Kong, K. Wu, and Y. Yu, “Distributional reinforcement learning for efficient exploration,” in International conference on Machine Learning . PMLR, 2019, pp. 4424–4434
2019
Earlier work this paper cites.
2019
Cited alongside, same era.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science robotics , vol. 5, no. 47, p. eabc5986, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Y. Ji, Z. Li, Y. Sun, X. B. Peng, S. Levine, G. Berseth, and K. Sreenath, “Hierarchical reinforcement learning for precise soccer shooting skills using a quadrupedal robot,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2022, pp. 1479–1486
2022
Later among the works it cites.
2022
Later among the works it cites.
Z. Fu, X. Cheng, and D. Pathak, “Learning a unified policy for whole-body control of manipulation and locomotion,” in Conference on Robot Learning (CoRL) , 2022
2022
Later among the works it cites.
Z. Fu, A. Kumar, A. Agarwal, H. Qi, J. Malik, and D. Pathak, “Coupling vision and proprioception for navigation of legged robots,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 17 273–17 283
2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
V. Makoviychuk, L. Wawrzyniak, Y. Guo, M. Lu, K. Storey, M. Macklin, D. Hoeller, N. Rudin, A. Allshire, A. Handa, et al. , “Isaac gym: High performance gpu-based physics simulation for robot learning,” in Conference on Neural Information Processing Systems (NeurIPS) , 2021
2021
Cited alongside, same era.
N. A. Urpí, S. Curi, and A. Krause, “Risk-averse offline reinforcement learning,” in International Conference on Learning Representations (ICLR) , 2021
2021
Cited alongside, same era.
N. A. Urpí, S. Curi, and A. Krause, “Risk-averse offline reinforcement learning,” in International Conference on Learning Representations (ICLR) , 2021. [Online]. Available: https://openreview.net/forum?id=TBIzh9b5eaz
2021
Cited alongside, same era.
Y. Ma, D. Jayaraman, and O. Bastani, “Conservative offline distributional reinforcement learning,” Advances in Neural Information Processing Systems , vol. 34, pp. 19 235–19 247, 2021
2021
Cited alongside, same era.
D. W. Nam, Y. Kim, and C. Y. Park, “Gmac: A distributional perspective on actor-critic framework,” in International Conference on Machine Learning . PMLR, 2021, pp. 7927–7936
2021
Cited alongside, same era.
N. Rudin, D. Hoeller, P. Reist, and M. Hutter, “Learning to walk in minutes using massively parallel deep reinforcement learning,” in Conference on Robot Learning (CoRL) , 2022
2022
Cited alongside, same era.
G. Margolis, G. Yang, K. Paigwar, T. Chen, and P. Agrawal, “Rapid locomotion via reinforcement learning,” in Robotics: Science and Systems , 2022
2022
Cited alongside, same era.
G. Margolis, G. Yang, K. Paigwar, T. Chen, and P. Agrawal, “Rapid locomotion via reinforcement learning,” in Robotics: Science and Systems , 2022
2022
Cited alongside, same era.
Later among the works it cites.
C. Bai, T. Xiao, Z. Zhu, L. Wang, F. Zhou, A. Garg, B. He, P. Liu, and Z. Wang, “Monotonic Quantile Network for Worst-Case Offline Reinforcement Learning,” IEEE Transactions on Neural Networks and Learning Systems , pp. 1–15, 2022
2022
Later among the works it cites.
A. Mavor-Parker, K. Young, C. Barry, and L. Griffin, “How to stay curious while avoiding noisy tvs using aleatoric uncertainty estimation,” in International Conference on Machine Learning . PMLR, 2022, pp. 15 220–15 240
2022
Later among the works it cites.
G. Bellegarda, Y. Chen, Z. Liu, and Q. Nguyen, “Robust high-speed running for quadruped robots via deep reinforcement learning,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2022, pp. 10 364–10 370
2022
Later among the works it cites.
I. M. A. Nahrendra, B. Yu, and H. Myung, “Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 5078–5084
2023
Closest in time.
R. Yang, G. Yang, and X. Wang, “Neural volumetric memory for visual locomotion control,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 1430–1440
2023
Closest in time.
S. Choi, G. Ji, J. Park, H. Kim, J. Mun, J. H. Lee, and J. Hwangbo, “Learning quadrupedal locomotion on deformable terrain,” Science Robotics , vol. 8, no. 74, p. eade2256, Jan. 2023
2023
Closest in time.
L. Wellhausen and M. Hutter, “ArtPlanner: Robust Legged Robot Navigation in the Field,” Field Robotics , vol. 3, no. 1, pp. 413–434, Jan. 2023
2023
Closest in time.
H. He, C. Bai, H. Lai, L. Wang, and W. Zhang, “Privileged knowledge distillation for sim-to-real policy generalization,” 2023
2023
Closest in time.
Y. Ma, F. Farshidian, and M. Hutter, “Learning arm-assisted fall damage reduction and recovery for legged mobile manipulators,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 12 149–12 155
2023
Closest in time.
G. B. Margolis and P. Agrawal, “Walk these ways: Tuning robot control for generalization with multiplicity of behavior,” in Conference on Robot Learning (CoRL) . PMLR, 2023, pp. 22–31
2023
Closest in time.
C. Li, M. Vlastelica, S. Blaes, J. Frey, F. Grimminger, and G. Martius, “Learning agile skills via adversarial imitation of rough partial demonstrations,” in Conference on Robot Learning (CoRL) . PMLR, 2023, pp. 342–352
2023
Closest in time.
2023
Closest in time.
P. Wu, A. Escontrela, D. Hafner, P. Abbeel, and K. Goldberg, “Daydreamer: World models for physical robot learning,” in Conference on Robot Learning (CoRL) . PMLR, 2023, pp. 2226–2240
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.