Fetching the paper…
Reading the bibliography…
In recent years, both reinforcement learning and learning-based control -- as well as the study of their safety, which is crucial for deployment in real-world robots -- have gained significant traction.
A. G. Barto, et al. , “Neuronlike adaptive elements that can solve difficult learning control problems,” IEEE Transactions on Systems, Man, and Cybernetics , vol. SMC-13, no. 5, pp. 834–846, 1983
1983
Earlier work this paper cites.
R. V. Florian, “Correct equations for the dynamics of the cart-pole system,” Center for Cognitive and Neural Studies (Coneural), Romania , 2007
2007
Earlier work this paper cites.
J. P. How, “Benchmarks [from the editor],” IEEE Control Systems Magazine , vol. 35, no. 1, pp. 6–7, 2015
2015
Earlier work this paper cites.
G. Brockman, et al. , “OpenAI Gym,” arXiv:1606.01540 [cs.LG] , 2016
2016
Earlier work this paper cites.
J. Leike, et al. , “AI safety gridworlds,” arXiv:1711.09883 [cs.LG] , 2017
2017
Earlier work this paper cites.
L. Pinto, et al. , “Robust adversarial reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning , 2017, vol. 70, pp. 2817–2826
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Schulman, et al. , “Proximal policy optimization algorithms,” arXiv:1707.06347 [cs.LG] , 2017
2017
Earlier work this paper cites.
P. Henderson, et al. , “Deep reinforcement learning that matters,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 32(1). AAAI Press, 2018
2018
Earlier work this paper cites.
S. Shah, et al. , “AirSim: High-fidelity visual and physical simulation for autonomous vehicles,” in Field and Service Robotics . Springer Int’l Publishing, 2018, pp. 621–635
2018
Earlier work this paper cites.
T. Haarnoja, et al. , “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in Proceedings of the 35th International Conference on Machine Learning , vol. 80, 2018, pp. 1861–1870
2018
Earlier work this paper cites.
G. Dalal, et al. , “Safe exploration in continuous action spaces,” arXiv:1801.08757 [cs.AI] , 2018
2018
Earlier work this paper cites.
K. P. Wabersich and M. N. Zeilinger, “Linear model predictive safety certification for learning-based control,” in 2018 IEEE Conference on Decision and Control (CDC) , 2018, pp. 7130–7135
2018
Cited alongside, same era.
L. Wang, et al. , “Safe learning of quadrotor dynamics using barrier certificates,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) , 2018, pp. 2460–2465
2018
Cited alongside, same era.
J. A. E. Andersson, et al. , “CasADi – A software framework for nonlinear optimization and optimal control,” Mathematical Programming Computation , vol. 11, no. 1, pp. 1–36, 2019
2019
Cited alongside, same era.
A. Ray, et al. , “Benchmarking safe exploration in deep reinforcement learning,” https://cdn.openai.com/safexp-short.pdf , 2019
2019
Cited alongside, same era.
B. Recht, “A tour of reinforcement learning: The view from continuous control,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 2, no. 1, pp. 253–279, 2019
L. Hewing, et al. , “Cautious model predictive control using gaussian process regression,” IEEE Transactions on Control Systems Technology , vol. 28, no. 6, pp. 2736–2743, 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
A. Taylor, et al. , “Learning for safety-critical control with control barrier functions,” in Proceedings of the 2nd Conference on Learning for Dynamics and Control , 2020, vol. 120, pp. 708–717
2020
Later among the works it cites.
2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
R. Julian, et al. , “Garage: A toolkit for reproducible reinforcement learning research,” https://github.com/rlworkgroup/garage , 2019
2019
Cited alongside, same era.
T. Wang, et al. , “Benchmarking model-based reinforcement learning,” arXiv:1907.02057 [cs.LG] , 2019
2019
Cited alongside, same era.
B. Ellenberger, “PyBullet Gymperium,” https://github.com/benelot/pybullet-gym , 2018–2019
2019
Cited alongside, same era.
A. Raffin, et al. , “Stable baselines3,” https://github.com/DLR-RM/stable-baselines3 , 2019
2019
Cited alongside, same era.
A. D. Ames, et al. , “Control barrier functions: Theory and applications,” in 2019 18th European Control Conference (ECC) , 2019, pp. 3420–3431
2019
Cited alongside, same era.
Y. Song, et al. , “Flightmare: A flexible quadrotor simulator,” in Proc. of the 4th Conference on Robot Learning , 2020
2020
Cited alongside, same era.
J. B. Rawlings, et al. , Model Predictive Control: Theory, Computation, and Design . Nob Hill Publishing, 2020, vol. 2nd
2020
Cited alongside, same era.
2021
Closest in time.
W. Tan-White, “Introducing intrinsic,” 2021. [Online]. Available: https://blog.x.company/introducing-intrinsic-1cf35b87651
2021
Closest in time.
E. Coumans and Y. Bai, “PyBullet, a Python module for physics simulation for games, robotics and machine learning,” http://pybullet.org , 2016–2021
2021
Closest in time.
2021
Closest in time.
J. Collins, et al. , “A review of physics simulators for robotic applications,” IEEE Access , vol. 9, pp. 51 416–51 431, 2021
2021
Closest in time.
J. Panerati, et al. , “Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,” in 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2021, pp. 7512–7519
2021
Closest in time.
J. R. Gardner, et al. , “GPyTorch: Blackbox matrix-matrix Gaussian process inference with GPU acceleration,” 2021
2021
Closest in time.
L. Brunke, et al. , “Safe learning in robotics: From learning-based control to safe reinforcement learning,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 5, no. 1, 2022. [Online]. Available: https://doi.org/10.1146/annurev-control-042920-020211
2022
Closest in time.