Fetching the paper…
Reading the bibliography…
Developing robust vision-guided controllers for quadrupedal robots in complex environments, with various obstacles, dynamical surroundings and uneven terrains, is very challenging.
K. Doya, “Reinforcement learning in continuous time and space,” Neural Comput. , vol. 12, no. 1, pp. 219–245, 2000. [Online]. Available: https://doi.org/10.1162/089976600300015961
2000
Earlier work this paper cites.
A. Gloye, M. Simon, A. Egorova, F. Wiesel, O. Tenchio, M. Schreiber, S. Behnke, and R. Rojas, “Predicting away robot control latency,” Freie Universität Berlin, Fachbereich Mathematik und Informatik, Serie B - Informatik B08-03, 2003
2003
Earlier work this paper cites.
K. V. Katsikopoulos and S. E. Engelbrecht, “Markov decision processes with delays and asynchronous cost collection,” IEEE Trans. Autom. Control. , vol. 48, no. 4, pp. 568–574, 2003. [Online]. Available: https://doi.org/10.1109/TAC.2003.809799
2003
Earlier work this paper cites.
N. Koenig and A. Howard, “Design and use paradigms for gazebo, an open-source multi-robot simulator,” in IEEE/RSJ International Conference on Intelligent Robots and Systems , Sendai, Japan, Sep 2004, pp. 2149–2154
2004
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2012 2012
2012
Earlier work this paper cites.
E. Coumans, “Bullet physics simulation,” in Special Interest Group on Computer Graphics and Interactive Techniques Conference, SIGGRAPH ’15, Los Angeles, CA, USA, August 9-13, 2015, Courses . ACM, 2015, p. 7:1
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
I. Mordatch, K. Lowrey, and E. Todorov, “Ensemble-cio: Full-body dynamic motion planning that transfers to physical humanoids,” in 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2015
2015
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” JMLR , vol. 17, pp. 39:1–39:40, 2016
2016
Earlier work this paper cites.
X. B. Peng, G. Berseth, K. Yin, and M. Van De Panne, “Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning,” ACM Transactions on Graphics (TOG) , vol. 36, no. 4, pp. 1–13, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in IROS 2017 . IEEE, 2017, pp. 23–30
2017
Earlier work this paper cites.
J. D. Carlo, P. M. Wensing, B. Katz, G. Bledt, and S. Kim, “Dynamic locomotion in the MIT cheetah 3 through convex model-predictive control,” in 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2018,
2018
Earlier work this paper cites.
A. Faust, K. Oslund, O. Ramirez, A. G. Francis, L. Tapia, M. Fiser, and J. Davidson, “PRM-RL: long-range robotic navigation tasks by combining reinforcement learning and sampling-based planning,” in IEEE International Conference on Robotics and Automation, ICRA 2018
2018
Earlier work this paper cites.
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen, “Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,” Int. J. Robotics Res. , vol. 37, no. 4-5, pp. 421–436, 2018
2018
Earlier work this paper cites.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in 2018 IEEE International Conference on Robotics and Automation, ICRA 2018 . IEEE, 2018, pp. 1–8
2018
Cited alongside, same era.
A. Sax, B. Emi, A. R. Zamir, L. J. Guibas, S. Savarese, and J. Malik, “Mid-level visual representations improve generalization and sample efficiency for learning visuomotor policies.” 2018
2018
Cited alongside, same era.
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y. Bai, D. Hafner, S. Bohez, and V. Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” in Robotics: Science and Systems XIV, 2018 , H. Kress-Gazit, S. S. Srinivasa, T. Howard, and N. Atanasov, Eds
2018
Cited alongside, same era.
Unitree, “A1: More dexterity, more posibility,” 2018. [Online]. Available: https://www.unitree.com/products/a1/
2018
Cited alongside, same era.
T. Xiao, E. Jang, D. Kalashnikov, S. Levine, J. Ibarz, K. Hausman, and A. Herzog, “Thinking while moving: Deep reinforcement learning with concurrent control,” in ICLR , 2020
2020
Later among the works it cites.
Z. Xie, P. Clary, J. Dao, P. Morais, J. Hurst, and M. Panne, “Learning locomotion skills for cassie: Iterative design and sim-to-real,” in Conference on Robot Learning . PMLR, 2020, pp. 317–329
2020
Later among the works it cites.
2020
Later among the works it cites.
E. Coumans and Y. Bai, “Pybullet, a python module for physics simulation for games, robotics and machine learning,” http://pybullet.org , 2016–2021
2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Carius, R. Ranftl, V. Koltun, and M. Hutter, “Trajectory optimization for legged robots with slipping motions,” IEEE Robotics and Automation Letters , vol. 4, no. 3, 2019
2019
Cited alongside, same era.
Y. Ding, A. Pandala, and H.-W. Park, “Real-time model predictive control for versatile dynamic motions in quadrupedal robots,” in ICRA . IEEE, 2019, pp. 8484–8490
2019
Cited alongside, same era.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, 2019
2019
Cited alongside, same era.
2020
Cited alongside, same era.
D. Jain, A. Iscen, and K. Caluwaerts, “From pixels to legs: Hierarchical learning of quadruped locomotion,” 2020
2020
Cited alongside, same era.
M. Laskin, K. Lee, A. Stooke, L. Pinto, P. Abbeel, and A. Srinivas, “Reinforcement learning with augmented data,” in NeurIPS , H. Larochelle, M. Ranzato, R. Hadsell, M. Balcan, and H. Lin, Eds., 2020. [Online]. Available: https://proceedings.neurips.cc/paper/2020/hash/e615c82aba461681ade82da2da38004a-Abstract.html
2020
Cited alongside, same era.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science robotics , vol. 5, no. 47, 2020
2020
Cited alongside, same era.
M. Li, Y. Wang, and D. Ramanan, “Towards streaming perception,” in ECCV , 2020
2020
Cited alongside, same era.
2021
Closest in time.
N. Hansen, H. Su, and X. Wang, “Stabilizing deep q-learning with convnets and vision transformers under data augmentation,” 2021
2021
Closest in time.
N. Hansen and X. Wang, “Generalization in reinforcement learning by soft data augmentation,” in ICRA , 2021
2021
Closest in time.
D. Hoeller, L. Wellhausen, F. Farshidian, and M. Hutter, “Learning a state representation and navigation in cluttered and dynamic environments,” IEEE Robotics and Automation Letters , 2021
2021
Closest in time.
A. Kumar, Z. Fu, D. Pathak, and J. Malik, “RMA: rapid motor adaptation for legged robots,” in Robotics: Science and Systems XVII, Virtual Event, July 12-16, 2021 , D. A. Shell, M. Toussaint, and M. A. Hsieh, Eds
2021
Closest in time.
Z. Li, X. Cheng, X. B. Peng, P. Abbeel, S. Levine, G. Berseth, and K. Sreenath, “Reinforcement learning for robust parameterized locomotion control of bipedal robots,” in 2021 IEEE International Conference on Robotics and Automation (ICRA)
2021
Closest in time.
2021
Closest in time.
M. Lutter, S. Mannor, J. Peters, D. Fox, and A. Garg, “Robust value iteration for continuous control tasks,” in RSS , July 2021
2021
Closest in time.
G. B. Margolis, T. Chen, K. Paigwar, X. Fu, D. Kim, S. bae Kim, and P. Agrawal, “Learning to jump from pixels,” in 5th Annual Conference on Robot Learning , 2021
2021
Closest in time.
D. Yarats, I. Kostrikov, and R. Fergus, “Image augmentation is all you need: Regularizing deep reinforcement learning from pixels,” in ICLR . OpenReview.net, 2021
2021
Closest in time.
R. Yang, M. Zhang, N. Hansen, H. Xu, and X. Wang, “Learning vision-guided quadrupedal locomotion end-to-end with cross-modal transformers,” in ICLR , 2022
2022
Closest in time.