Fetching the paper…
Reading the bibliography…
Deep Neural Network-based systems are now the state-of-the-art in many robotics tasks, but their application in safety-critical domains remains dangerous without formal guarantees on network robustness.
1910
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International conference on machine learning , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
A. G. Barto, R. S. Sutton, and C. W. Anderson, “Neuronlike adaptive elements that can solve difficult learning control problems,” IEEE transactions on systems, man, and cybernetics , no. 5, pp. 834–846, 1983
1983
Earlier work this paper cites.
M. L. Littman, “Markov games as a framework for multi-agent reinforcement learning,” in Machine learning proceedings 1994 . Elsevier, 1994, pp. 157–163
1994
Earlier work this paper cites.
M. Heger, “Consideration of risk in reinforcement learning,” in Machine Learning Proceedings 1994 , W. W. Cohen and H. Hirsh, Eds. San Francisco (CA): Morgan Kaufmann, 1994, pp. 105 – 111
1994
Earlier work this paper cites.
B. W. Mott, T. Team et al. , “Stella: a multiplatform atari 2600 vcs emulator,” 1995
1995
Earlier work this paper cites.
W. Uther and M. Veloso, “Adversarial reinforcement learning,” In Proceedings of the AAAI Fall Symposium on Model Directed Autonomous Systems, Tech. Rep., 1997
1997
Earlier work this paper cites.
S. Boyd and L. Vandenberghe, Convex optimization . Cambridge university press, 2004
2004
Earlier work this paper cites.
J. Morimoto and K. Doya, “Robust reinforcement learning,” Neural computation , vol. 17, no. 2, pp. 335–359, 2005
2005
Earlier work this paper cites.
P. Geibel, “Risk-sensitive approaches for reinforcement learning,” Ph.D. dissertation, University of Osnabrück, 2006
2006
Earlier work this paper cites.
J. P. van den Berg, S. J. Guy, M. C. Lin, and D. Manocha, “Reciprocal n-body collision avoidance,” in International Symposium on Robotics Research (ISRR) , 2009
2009
Earlier work this paper cites.
J. Snoek, H. Larochelle, and R. P. Adams, “Practical bayesian optimization of machine learning algorithms,” in Advances in Neural Information Processing Systems (NeurIPS) 25 , F. Pereira, C. J. C. Burges, L. Bottou, and K. Q. Weinberger, Eds. Curran Associates, Inc., 2012, pp. 2951–2959
2012
Earlier work this paper cites.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling, “The arcade learning environment: An evaluation platform for general agents,” Journal of Artificial Intelligence Research , vol. 47, pp. 253–279, jun 2013
2013
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in International Conference on Learning Representations (ICLR) , 2014
2014
Earlier work this paper cites.
I. Goodfellow, J. Shlens, and C. Szegedy, “Explaining and harnessing adversarial examples,” in International Conference on Learning Representations (ICLR) , 2015
2015
Earlier work this paper cites.
J. García and F. Fernández, “A comprehensive survey on safe reinforcement learning,” Journal of Machine Learning Research , vol. 16, pp. 1437–1480, 2015
2015
Earlier work this paper cites.
A. Tamar, “Risk-sensitive and efficient reinforcement learning algorithms,” Ph.D. dissertation, Technion - Israel Institute of Technology, Faculty of Electrical Engineering, 2015
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” in Nature . Nature Publishing Group, a division of Macmillan Publishers Limited., 2015, vol. 518
2015
Earlier work this paper cites.
M. Sharif, S. Bhagavatula, L. Bauer, and M. K. Reiter, “Accessorize to a crime: Real and stealthy attacks on state-of-the-art face recognition,” in Proceedings of the 2016 ACM SIGSAC Conference on Computer and Communications Security , 2016, pp. 1528–1540
2016
Earlier work this paper cites.
N. Papernot, P. McDaniel, X. Wu, S. Jha, and A. Swami, “Distillation as a defense to adversarial perturbations against deep neural networks,” in 2016 IEEE Symposium on Security and Privacy (SP) , May 2016, pp. 582–597
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Earlier work this paper cites.
A. S. Bandeira, “A note on probably certifiably correct algorithms,” Comptes Rendus Mathematique , vol. 354, no. 3, pp. 329–333, 2016
2016
Earlier work this paper cites.
S. Gu, E. Holly, T. Lillicrap, and S. Levine, “Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates,” in 2017 IEEE International Conference on Robotics and Automation (ICRA) , May 2017
2017
Earlier work this paper cites.
A. Kurakin, I. J. Goodfellow, and S. Bengio, “Adversarial examples in the physical world,” in International Conference on Learning Representation (ICLR) (Workshop) , 2017
2017
Earlier work this paper cites.
A. Mandlekar, Y. Zhu, A. Garg, L. Fei-Fei, and S. Savarese, “Adversarially robust policy learning: Active construction of physically-plausible perturbations,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2017, pp. 3932–3939
2017
Earlier work this paper cites.
A. Rajeswaran, S. Ghotra, B. Ravindran, and S. Levine, “Epopt: Learning robust neural network policies using model ensembles,” in International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning (ICML) , ser. Proceedings of Machine Learning Research, D. Precup and Y. W. Teh, Eds., vol. 70. International Convention Centre, Sydney, Australia: PMLR, 06–11 Aug 2017, pp. 2817–2826
2017
Cited alongside, same era.
R. Ehlers, “Formal verification of piece-wise linear feed-forward neural networks,” in ATVA , 2017
2017
Cited alongside, same era.
G. Katz, C. W. Barrett, D. L. Dill, K. Julian, and M. J. Kochenderfer, “Reluplex: An efficient SMT solver for verifying deep neural networks,” in Computer Aided Verification - 29th International Conference, CAV 2017, Heidelberg, Germany, July 24-28, 2017, Proceedings, Part I , 2017, pp. 97–117
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu, “Towards deep learning models resistant to adversarial attacks,” in International Conference on Learning Representations (ICLR) , 2018
2018
Later among the works it cites.
F. Tramèr, A. Kurakin, N. Papernot, I. Goodfellow, D. Boneh, and P. McDaniel, “Ensemble adversarial training: Attacks and defenses,” in International Conference on Learning Representations (ICLR) , 2018
2018
Later among the works it cites.
W. Xu, D. Evans, and Y. Qi, “Feature squeezing: Detecting adversarial examples in deep neural networks,” in Network and Distributed Systems Security Symposium (NDSS) . The Internet Society, 2018
2018
Later among the works it cites.
A. Athalye, N. Carlini, and D. Wagner, “Obfuscated gradients give a false sense of security: Circumventing defenses to adversarial examples,” in Proceedings of the 35th International Conference on Machine Learning (ICML) , ser. Proceedings of Machine Learning Research, J. Dy and A. Krause, Eds., vol. 80. Stockholmsmässan, Stockholm Sweden: PMLR, 10–15 Jul 2018, pp. 274–283
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
X. Huang, M. Kwiatkowska, S. Wang, and M. Wu, “Safety verification of deep neural networks,” in Computer Aided Verification , R. Majumdar and V. Kunčak, Eds. Cham: Springer International Publishing, 2017, pp. 3–29
2017
Cited alongside, same era.
2017
Cited alongside, same era.
S. Huang, N. Papernot, I. Goodfellow, Y. Duan, and P. Abbeel, “Adversarial attacks on neural network policies,” 2017
2017
Cited alongside, same era.
V. Behzadan and A. Munir, “Vulnerability of deep reinforcement learning to policy induction attacks,” in International Conference on Machine Learning and Data Mining in Pattern Recognition (MLDM) . Springer, 2017, pp. 262–275
2017
Cited alongside, same era.
A. Kurakin, I. J. Goodfellow, and S. Bengio, “Adversarial machine learning at scale,” in International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
J. Kos and D. Song, “Delving into adversarial attacks on deep policies,” in International Conference on Learning Representations (ICLR) (Workshop) , 2017
2017
Cited alongside, same era.
N. Carlini and D. Wagner, “Adversarial examples are not easily detected: Bypassing ten detection methods,” in Proceedings of the 10th ACM Workshop on Artificial Intelligence and Security , ser. AISec ’17. New York, NY, USA: ACM, 2017, pp. 3–14
2017
Cited alongside, same era.
W. He, J. Wei, X. Chen, N. Carlini, and D. Song, “Adversarial example defenses: Ensembles of weak defenses are not strong,” in Proceedings of the 11th USENIX Conference on Offensive Technologies , ser. WOOT’17. Berkeley, CA, USA: USENIX Association, 2017, pp. 15–15
2017
Cited alongside, same era.
2018
Later among the works it cites.
J. Uesato, B. O’Donoghue, P. Kohli, and A. van den Oord, “Adversarial risk and the dangers of evaluating against weak attacks,” in Proceedings of the 35th International Conference on Machine Learning (ICML) , ser. Proceedings of Machine Learning Research, J. Dy and A. Krause, Eds., vol. 80. Stockholmsmässan, Stockholm Sweden: PMLR, 10–15 Jul 2018, pp. 5025–5034
2018
Later among the works it cites.
2018
Later among the works it cites.
A. Hill, A. Raffin, M. Ernestus, A. Gleave, A. Kanervisto, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Stable baselines,” 2018. [Online]. Available: https://github.com/hill-a/stable-baselines
2018
Later among the works it cites.
A. Raffin, “Rl baselines zoo,” https://github.com/araffin/rl-baselines-zoo , 2018
2018
Later among the works it cites.
T. Fan, X. Cheng, J. Pan, P. Long, W. Liu, R. Yang, and D. Manocha, “Getting robots unfrozen and unlost in dense pedestrian crowds,” IEEE Robotics and Automation Letters , vol. 4, no. 2, pp. 1178–1185, 2019
2019
Later among the works it cites.
X. Yuan, P. He, Q. Zhu, R. R. Bhat, and X. Li, “Adversarial examples: Attacks and defenses for deep learning,” IEEE transactions on neural networks and learning systems , 2019
2019
Later among the works it cites.
Tencent Keen Security Lab, “Experimental security research of Tesla Autopilot,” 03 2019. [Online]. Available: https://keenlab.tencent.com/en/whitepapers/\Experimental_Security_Research_of_Tesla_Autopilot.pdf
2019
Later among the works it cites.
V. Tjeng, K. Y. Xiao, and R. Tedrake, “Evaluating robustness of neural networks with mixed integer programming,” in International Conference on Learning Representations (ICLR) , 2019
2019
Later among the works it cites.
G. Singh, T. Gehr, M. Püschel, and M. Vechev, “An abstract domain for certifying neural networks,” Proceedings of the ACM on Programming Languages , vol. 3, no. POPL, pp. 1–30, 2019
2019
Later among the works it cites.
H. Salman, G. Yang, H. Zhang, C.-J. Hsieh, and P. Zhang, “A convex relaxation barrier to tight robustness verification of neural networks,” in Advances in Neural Information Processing Systems , 2019, pp. 9832–9842
2019
Later among the works it cites.
M. Mirman, M. Fischer, and M. Vechev, “Distilled agent DQN for provable adversarial robustness,” 2019. [Online]. Available: https://openreview.net/forum?id=ryeAy3AqYm
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
E. Wong, F. R. Schmidt, and J. Z. Kolter, “Wasserstein adversarial examples via projected sinkhorn iterations,” vol. 97, 2019
2019
Later among the works it cites.
Y.-S. Wang, T.-W. Weng, and L. Daniel, “Verification of neural network control policy under persistent adversarial perturbation,” NeurIPS Workshop on Safety and Robustness in Decision Making , 2019
2019
Later among the works it cites.
2020
Closest in time.
C.-H. H. Yang, J. Qi, P.-Y. Chen, Y. Ouyang, I.-T. D. Hung, C.-H. Lee, and X. Ma, “Enhanced adversarial strategically-timed attacks against deep reinforcement learning,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2020, pp. 3407–3411
2020
Closest in time.
A. Gleave, M. Dennis, N. Kant, C. Wild, S. Levine, and S. Russell, “Adversarial policies: Attacking deep reinforcement learning,” 2020
2020
Closest in time.
M. Everett and J. How, “Gym: Collision avoidance,” 2020. [Online]. Available: https://github.com/mit-acl/gym-collision-avoidance
2020
Closest in time.
H. Zhang, H. Chen, C. Xiao, S. Gowal, R. Stanforth, B. Li, D. Boning, and C.-J. Hsieh, “Crown-ibp: Towards stable and efficient training of verifiably robust neural networks,” 2020. [Online]. Available: https://github.com/huanzhang12/CROWN-IBP
2020
Closest in time.
H. Yang, J. Shi, and L. Carlone, “Teaser: Fast and certifiable point cloud registration,” 2020. [Online]. Available: https://github.com/MIT-SPARK/TEASER-plusplus
2020
Closest in time.