Fetching the paper…
Reading the bibliography…
Current robot platforms available for research are either very expensive or unable to handle the abuse of exploratory controls in reinforcement learning.
A. H. Schoen, Infinite periodic minimal surfaces without self-intersections . National Aeronautics and Space Administration, 1970
1970
Earlier work this paper cites.
G. Bradski, “The OpenCV Library,” Dr. Dobb’s Journal of Software Tools , 2000
2000
Earlier work this paper cites.
P. Holoborodko, “Smooth noise robust differentiators,” http://www.holoborodko.com/pavel/numerical-methods/numerical-derivative/smooth-low-noise-differentiators/, 2008
2008
Earlier work this paper cites.
2016
Earlier work this paper cites.
J. Schulman, P. Moritz, S. Levine, M. Jordan, and P. Abbeel, “High-dimensional continuous control using generalized advantage estimation,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2016
2016
Earlier work this paper cites.
S. Garrido-Jurado, R. Munoz-Salinas, F. J. Madrid-Cuevas, and R. Medina-Carnicer, “Generation of fiducial marker dictionaries using mixed integer linear programming,” Pattern Recognition , vol. 51, pp. 481–491, 2016
2016
Earlier work this paper cites.
L. Paull, J. Tani, H. Ahn, J. Alonso-Mora, L. Carlone, M. Cap, Y. F. Chen, C. Choi, J. Dusek, Y. Fang, et al. , “Duckietown: an open, inexpensive and flexible platform for autonomy education and research,” in 2017 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2017, pp. 1497–1504
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Gupta, A. Murali, D. P. Gandhi, and L. Pinto, “Robot learning in homes: Improving generalization and reducing dataset bias,” in Advances in Neural Information Processing Systems , 2018, pp. 9094–9104
2018
Cited alongside, same era.
F. J. Romero-Ramirez, R. Muñoz-Salinas, and R. Medina-Carnicer, “Speeded up detection of squared fiducial markers,” Image and vision Computing , vol. 76, pp. 38–47, 2018
2018
Cited alongside, same era.
S. Fujimoto, H. Hoof, and D. Meger, “Addressing function approximation error in actor-critic methods,” in International Conference on Machine Learning , 2018, pp. 1587–1596
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International Conference on Machine Learning , 2018, pp. 1861–1870
2018
Cited alongside, same era.
D. V. Gealy, S. McKinley, B. Yi, P. Wu, P. R. Downey, G. Balke, A. Zhao, M. Guo, R. Thomasson, A. Sinclair, et al. , “Quasi-direct drive for low-cost compliant robotic manipulation,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 437–443
2019
Later among the works it cites.
B. Yang, J. Zhang, D. Jayaraman, and S. Levine, “REPLAB: A reproducible low-cost arm benchmark platform for robotic learning,” ICRA , 2019
2019
Later among the works it cites.
M. Janner, J. Fu, M. Zhang, and S. Levine, “When to trust your model: model-based policy optimization,” in Advances in Neural Information Processing Systems , 2019, pp. 12 519–12 530
2019
Later among the works it cites.
F. Grimminger, A. Meduri, M. Khadiv, J. Viereck, M. Wüthrich, M. Naveau, V. Berenz, S. Heim, F. Widmaier, T. Flayols, et al. , “An open torque-controlled modular robot architecture for legged locomotion research,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 3650–3657, 2020
2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Coumans and Y. Bai, “PyBullet, a Python module for physics simulation for games, robotics and machine learning,” http://pybullet.org , 2016–2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
B. Katz, J. Di Carlo, and S. Kim, “Mini cheetah: A platform for pushing the limits of dynamic quadruped control,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 6295–6301
2019
Cited alongside, same era.
N. Kau, A. Schultz, N. Ferrante, and P. Slade, “Stanford doggo: An open-source, quasi-direct-drive quadruped,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 6309–6315
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Closest in time.
M. Ahn, H. Zhu, K. Hartikainen, H. Ponte, A. Gupta, S. Levine, and V. Kumar, “ROBEL: Robotics benchmarks for learning with low-cost robots,” in Conference on Robot Learning . PMLR, 2020, pp. 1300–1313
2020
Closest in time.
2020
Closest in time.
R. Boney, J. Kannala, and A. Ilin, “Regularizing model-based planning with energy-based models,” in Conference on Robot Learning . PMLR, 2020, pp. 182–191
2020
Closest in time.
X. Chen, C. Wang, Z. Zhou, and K. W. Ross, “Randomized ensembled double q-learning: Learning fast without a model,” in International Conference on Learning Representations , 2021
2021
Closest in time.