Fetching the paper…
Reading the bibliography…
Real-world autonomous missions often require rich interaction with nearby objects, such as doors or switches, along with effective navigation.
M. H. Raibert, “Trotting, pacing and bounding by a quadruped robot,” Journal of biomechanics , vol. 23, pp. 79–98, 1990
1990
Earlier work this paper cites.
R. S. Sutton, D. Precup, and S. Singh, “Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,” Artificial intelligence , vol. 112, no. 1-2, pp. 181–211, 1999
1999
Earlier work this paper cites.
L. Sentis and O. Khatib, “A whole-body control framework for humanoids operating in human environments,” in Proceedings 2006 IEEE International Conference on Robotics and Automation, 2006. ICRA 2006. IEEE, 2006, pp. 2641–2648
2006
Earlier work this paper cites.
G. Konidaris, S. Kuindersma, R. Grupen, and A. Barto, “Autonomous skill acquisition on a mobile manipulator,” in Twenty-Fifth AAAI Conference on Artificial Intelligence , 2011
2011
Earlier work this paper cites.
M. Hutter, C. Gehring, D. Jud, A. Lauber, C. D. Bellicoso, V. Tsounis, J. Hwangbo, K. Bodie, P. Fankhauser, M. Bloesch, et al. , “Anymal-a highly mobile and dynamic quadrupedal robot,” in 2016 IEEE/RSJ international conference on intelligent robots and systems (IROS) . IEEE, 2016, pp. 38–44
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, 2017
2017
Earlier work this paper cites.
P.-L. Bacon, J. Harb, and D. Precup, “The option-critic architecture,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 31, no. 1, 2017
2017
Earlier work this paper cites.
X. B. Peng, G. Berseth, K. Yin, and M. Van De Panne, “Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning,” ACM Transactions on Graphics (TOG) , vol. 36, no. 4, pp. 1–13, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
T. Apgar, P. Clary, K. Green, A. Fern, and J. W. Hurst, “Fast online trajectory optimization for the bipedal robot cassie.” in Robotics: Science and Systems , vol. 101, 2018, p. 14
2018
Earlier work this paper cites.
G. Bledt, M. J. Powell, B. Katz, J. Di Carlo, P. M. Wensing, and S. Kim, “Mit cheetah 3: Design and control of a robust, dynamic quadruped robot,” in IROS . IEEE, 2018, pp. 2245–2252
2018
Earlier work this paper cites.
J. Di Carlo, P. M. Wensing, B. Katz, G. Bledt, and S. Kim, “Dynamic locomotion in the mit cheetah 3 through convex model-predictive control,” in IROS . IEEE, 2018, pp. 1–9
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
T. Johannink, S. Bahl, A. Nair, J. Luo, A. Kumar, M. Loskyll, J. A. Ojea, E. Solowjow, and S. Levine, “Residual reinforcement learning for robot control,” in 2019 ICRA . IEEE, 2019, pp. 6023–6029
2019
Cited alongside, same era.
2019
Cited alongside, same era.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, p. eaau5872, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2020
Later among the works it cites.
T. Li, N. Lambert, R. Calandra, F. Meier, and A. Rai, “Learning generalizable locomotion skills with hierarchical reinforcement learning,” in 2020 IEEE ICRA . IEEE, 2020, pp. 413–419
2020
Later among the works it cites.
S. Zimmermann, R. Poranne, and S. Coros, “Go fetch!-dynamic grasps using boston dynamics spot with external robotic arm,” in 2021 IEEE ICRA . IEEE, 2021, pp. 4488–4494
2021
Later among the works it cites.
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Jain, A. Iscen, and K. Caluwaerts, “Hierarchical reinforcement learning for quadruped locomotion,” in 2019 IEEE/RSJ IROS . IEEE, 2019, pp. 7551–7557
2019
Cited alongside, same era.
X. B. Peng, M. Chang, G. Zhang, P. Abbeel, and S. Levine, “Mcp: Learning composable hierarchical control with multiplicative compositional policies,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
C. Yang, K. Yuan, Q. Zhu, W. Yu, and Z. Li, “Multi-expert learning of adaptive legged locomotion,” Science Robotics , vol. 5, no. 49, p. eabb2174, 2020
2020
Cited alongside, same era.
C. Li, F. Xia, R. Martin-Martin, and S. Savarese, “Hrl4in: Hierarchical reinforcement learning for interactive navigation with mobile manipulators,” in Conference on Robot Learning . PMLR, 2020, pp. 603–616
2020
Cited alongside, same era.
2020
Cited alongside, same era.
W. Yu, J. Tan, Y. Bai, E. Coumans, and S. Ha, “Learning fast adaptation with meta strategy optimization,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 2950–2957, 2020
2020
Cited alongside, same era.
2021
Later among the works it cites.
M. Sorokin, W. Yu, S. Ha, and C. K. Liu, “Learning human search behavior from egocentric visual inputs,” in Computer Graphics Forum , vol. 40, no. 2. Wiley Online Library, 2021, pp. 389–398
2021
Later among the works it cites.
K.-H. Zeng, L. Weihs, A. Farhadi, and R. Mottaghi, “Pushing it out of the way: Interactive visual navigation,” in Proceedings of the IEEE/CVF CVPR , 2021, pp. 9868–9877
2021
Later among the works it cites.
S. Pateria, B. Subagdja, A.-h. Tan, and C. Quek, “Hierarchical reinforcement learning: A comprehensive survey,” ACM Computing Surveys (CSUR) , vol. 54, no. 5, pp. 1–35, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2022
Closest in time.
S. Kim, M. Sorokin, J. Lee, and S. Ha, “Human motion control of quadrupedal robots using deep reinforcement learning,” Robotics Science and Systems , 2022
2022
Closest in time.
T. Miki, J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning robust perceptive locomotion for quadrupedal robots in the wild,” Science Robotics , vol. 7, no. 62, p. eabk2822, 2022
2022
Closest in time.
Z. Fu, A. Kumar, A. Agarwal, H. Qi, J. Malik, and D. Pathak, “Coupling vision and proprioception for navigation of legged robots,” in Proceedings of the IEEE/CVF CVPR , 2022, pp. 17 273–17 283
2022
Closest in time.
N. Rudin, D. Hoeller, P. Reist, and M. Hutter, “Learning to walk in minutes using massively parallel deep reinforcement learning,” in Conference on Robot Learning . PMLR, 2022, pp. 91–100
2022
Closest in time.
L. Smith, J. C. Kew, X. B. Peng, S. Ha, J. Tan, and S. Levine, “Legged robots that keep on learning: Fine-tuning locomotion policies in the real world,” in 2022 ICRA . IEEE, 2022, pp. 1593–1599
2022
Closest in time.
K. N. Kumar, I. Essa, and S. Ha, “Graph-based cluttered scene generation and interactive exploration using deep reinforcement learning,” in 2022 ICRA . IEEE, 2022, pp. 7521–7527
2022
Closest in time.