Fetching the paper…
Reading the bibliography…
Learning-based methods could provide solutions to many of the long-standing challenges in control.
B. Widrow, “Pattern-recognizing control systems,” Compurter and Information Sciences
1964
Earlier work this paper cites.
K.-S. Fu, “Learning control systems–review and outlook,” IEEE transactions on Automatic Control
1970
Earlier work this paper cites.
K. Fukushima and S. Miyake, “Neocognitron: A self-organizing neural network model for a mechanism of visual pattern recognition,” in Competition and cooperation in neural nets
1982
Earlier work this paper cites.
D. Psaltis, A. Sideris, and A. A. Yamamura, “A multilayered neural network controller,” IEEE control systems magazine
1988
Earlier work this paper cites.
W. Li and J.-J. E. Slotine, “Neural network control of unknown nonlinear systems,” in 1989 American Control Conference
1989
Earlier work this paper cites.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” tech. rep., CARNEGIE-MELLON UNIV PITTSBURGH PA ARTIFICIAL INTELLIGENCE AND PSYCHOLOGY …, 1989
1989
Earlier work this paper cites.
K. S. Narendra, “Adaptive control using neural networks,” Neural networks for control
1991
Earlier work this paper cites.
B. W. Mott, T. Team, et al
1995
Earlier work this paper cites.
K. Tanaka, “An approach to stability criteria of neural-network control systems,” IEEE Transactions on Neural Networks
1996
Earlier work this paper cites.
W. Uther and M. Veloso, “Adversarial reinforcement learning,” tech. rep., In Proceedings of the AAAI Fall Symposium on Model Directed Autonomous Systems, 1997
1997
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation
1997
Earlier work this paper cites.
C. J. Tomlin, J. Lygeros, and S. S. Sastry, “A game theoretic approach to controller design for hybrid systems,” Proceedings of the IEEE
2000
Earlier work this paper cites.
A. Papachristodoulou and S. Prajna, “On the construction of lyapunov functions using the sum of squares decomposition,” in Proceedings of the 41st IEEE Conference on Decision and Control, 2002
2002
Earlier work this paper cites.
J. Morimoto and K. Doya, “Robust reinforcement learning,” Neural computation
2005
Earlier work this paper cites.
R. Zhan and J. Wan, “Neural network-aided adaptive unscented kalman filter for nonlinear state estimation,” IEEE Signal Processing Letters
2006
Earlier work this paper cites.
N. Noroozi, P. Karimaghaee, F. Safaei, and H. Javadi, “Generation of lyapunov functions by neural networks,” in Proceedings of the World Congress on Engineering
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in 2009 IEEE conference on computer vision and pattern recognition
2009
Earlier work this paper cites.
Springer Science & Business Media, 2009
P. Tabuada, Verification and control of hybrid systems: a symbolic approach · 2009
Earlier work this paper cites.
Addison-Wesley Professional, 2010
J. Sanders and E. Kandrot, CUDA by example: an introduction to general-purpose GPU programming · 2010
Earlier work this paper cites.
G. Frehse, C. Le Guernic, A. Donzé, S. Cotton, R. Ray, O. Lebeltel, R. Ripado, A. Girard, T. Dang, and O. Maler, “Spaceex: Scalable verification of hybrid systems,” in International Conference on Computer Aided Verification
2011
Earlier work this paper cites.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling, “The arcade learning environment: An evaluation platform for general agents,” Journal of Artificial Intelligence Research
2013
Earlier work this paper cites.
X. Chen, E. Ábrahám, and S. Sankaranarayanan, “Flow*: An analyzer for non-linear hybrid systems,” in International Conference on Computer Aided Verification
2013
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in International Conference on Learning Representations (ICLR)
2014
Earlier work this paper cites.
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell, “Caffe: Convolutional architecture for fast feature embedding,” in Proceedings of the 22nd ACM international conference on Multimedia
2014
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,” The journal of machine learning research
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. C. Courville, and Y. Bengio, “Generative adversarial nets,” in NIPS
2014
Earlier work this paper cites.
S. Levine and P. Abbeel, “Learning neural network policies with guided policy search under unknown dynamics.,” in NIPS
2014
Earlier work this paper cites.
X. Zheng, C. Julien, M. Kim, and S. Khurshid, “Perceptions on the state of the art in verification and validation in cyber-physical systems,” IEEE Systems Journal
2015
Earlier work this paper cites.
M. Althoff, “An introduction to cora 2015,” in Proc. of the Workshop on Applied Verification for Continuous and Hybrid Systems
2015
Earlier work this paper cites.
P. S. Duggirala, S. Mitra, M. Viswanathan, and M. Potok, “C2e2: A verification tool for stateflow models,” in International Conference on Tools and Algorithms for the Construction and Analysis of Systems
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” in Nature
2015
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Earlier work this paper cites.
C. Fan, B. Qi, S. Mitra, M. Viswanathan, and P. S. Duggirala, “Automatic reachability analysis for nonlinear hybrid models with c2e2,” in International Conference on Computer Aided Verification
2016
Earlier work this paper cites.
A. D. Ames, X. Xu, J. W. Grizzle, and P. Tabuada, “Control barrier function based quadratic programs for safety critical systems,” IEEE Transactions on Automatic Control
2016
Earlier work this paper cites.
S. Diamond and S. Boyd, “Cvxpy: A python-embedded modeling language for convex optimization,” The Journal of Machine Learning Research
2016
Earlier work this paper cites.
F. Lamnabhi-Lagarrigue, A. Annaswamy, S. Engell, A. Isaksson, P. Khargonekar, R. M. Murray, H. Nijmeijer, T. Samad, D. Tilbury, and P. Van den Hof, “Systems & control for the future of humanity, research agenda: Current and future roles, impact and grand challenges,” Annual Reviews in Control
2017
Cited alongside, same era.
S. Huang, N. Papernot, I. Goodfellow, Y. Duan, and P. Abbeel, “Adversarial attacks on neural network policies,” 2017
2017
Cited alongside, same era.
V. Behzadan and A. Munir, “Vulnerability of deep reinforcement learning to policy induction attacks,” in International Conference on Machine Learning and Data Mining in Pattern Recognition (MLDM)
2017
Cited alongside, same era.
A. Mandlekar, Y. Zhu, A. Garg, L. Fei-Fei, and S. Savarese, “Adversarially robust policy learning: Active construction of physically-plausible perturbations,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
2017
Cited alongside, same era.
V. Tjeng, K. Y. Xiao, and R. Tedrake, “Evaluating robustness of neural networks with mixed integer programming,” in International Conference on Learning Representations (ICLR)
2019
Later among the works it cites.
S. Dutta, X. Chen, and S. Sankaranarayanan, “Reachability analysis for neural feedback systems using regressive polynomial rule inference,” in Proceedings of the 22nd ACM International Conference on Hybrid Systems: Computation and Control
2019
Later among the works it cites.
C. Huang, J. Fan, W. Li, X. Chen, and Q. Zhu, “Reachnn: Reachability analysis of neural-network controlled systems,” ACM Transactions on Embedded Computing Systems (TECS)
2019
Later among the works it cites.
R. Ivanov, J. Weimer, R. Alur, G. J. Pappas, and I. Lee, “Verisig: verifying safety properties of hybrid systems with neural network controllers,” in Proceedings of the 22nd ACM International Conference on Hybrid Systems: Computation and Control
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Rajeswaran, S. Ghotra, B. Ravindran, and S. Levine, “Epopt: Learning robust neural network policies using model ensembles,” in International Conference on Learning Representations (ICLR)
2017
Cited alongside, same era.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning (ICML)
2017
Cited alongside, same era.
N. P. Jouppi, C. Young, N. Patil, D. Patterson, G. Agrawal, R. Bajwa, S. Bates, S. Bhatia, N. Boden, A. Borchers, et al
2017
Cited alongside, same era.
2017
Cited alongside, same era.
S. Bansal, M. Chen, S. Herbert, and C. J. Tomlin, “Hamilton-jacobi reachability: A brief overview and recent advances,” in 2017 IEEE 56th Annual Conference on Decision and Control (CDC)
2017
Cited alongside, same era.
R. Ehlers, “Formal verification of piece-wise linear feed-forward neural networks,” in ATVA
2017
Cited alongside, same era.
G. Katz, C. W. Barrett, D. L. Dill, K. Julian, and M. J. Kochenderfer, “Reluplex: An efficient SMT solver for verifying deep neural networks,” in Computer Aided Verification - 29th International Conference, CAV 2017, Heidelberg, Germany, July 24-28, 2017, Proceedings, Part I
2017
Cited alongside, same era.
X. Huang, M. Kwiatkowska, S. Wang, and M. Wu, “Safety verification of deep neural networks,” in Computer Aided Verification
2017
Cited alongside, same era.
G. Yang, G. Qian, P. Lv, and H. Li, “Efficient verification of control systems with neural network controllers,” in Proceedings of the 3rd International Conference on Vision, Image and Signal Processing
2019
Later among the works it cites.
H. Salman, G. Yang, H. Zhang, C.-J. Hsieh, and P. Zhang, “A convex relaxation barrier to tight robustness verification of neural networks,” in Advances in Neural Information Processing Systems
2019
Later among the works it cites.
G. Singh, R. Ganvir, M. Püschel, and M. Vechev, “Beyond the single neuron convex barrier for neural network certification,” in Advances in Neural Information Processing Systems
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Fazlyab, M. Morari, and G. J. Pappas, “Probabilistic verification and reachability analysis of neural networks via semidefinite programming,” in 2019 IEEE 58th Conference on Decision and Control (CDC)
2019
Later among the works it cites.
P. Christodoulou, “Soft actor-critic for discrete action settings,” arXiv preprint arXiv:1910.07207
2019
Later among the works it cites.
C.-H. H. Yang, J. Qi, P.-Y. Chen, Y. Ouyang, I.-T. D. Hung, C.-H. Lee, and X. Ma, “Enhanced adversarial strategically-timed attacks against deep reinforcement learning,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2020
Later among the works it cites.
A. Gleave, M. Dennis, N. Kant, C. Wild, S. Levine, and S. Russell, “Adversarial policies: Attacking deep reinforcement learning,” 2020
2020
Later among the works it cites.
M. Everett, G. Habibi, and J. P. How, “Robustness analysis of neural networks via efficient partitioning with applications in control systems,” IEEE Control Systems Letters
2020
Later among the works it cites.
M. Cranmer, S. Greydanus, S. Hoyer, P. Battaglia, D. Spergel, and S. Ho, “Lagrangian neural networks,” in ICLR 2020 Workshop on Integration of Deep Neural Models and Differential Equations
2020
Later among the works it cites.
2020
Later among the works it cites.
A. Tagliabue, A. Paris, S. Kim, R. Kubicek, S. Bergbreiter, and J. P. How, “Touch the wind: Simultaneous airflow, drag and interaction sensing on a multirotor,” in 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
2020
Later among the works it cites.
M. Fazlyab, M. Morari, and G. J. Pappas, “Safety verification and robustness analysis of neural networks via quadratic constraints and semidefinite programming,” IEEE Transactions on Automatic Control
2020
Later among the works it cites.
K. Xu, Z. Shi, H. Zhang, Y. Wang, K.-W. Chang, M. Huang, B. Kailkhura, X. Lin, and C.-J. Hsieh, “Automatic perturbation analysis for scalable certified robustness and beyond,” Advances in Neural Information Processing Systems
2020
Later among the works it cites.
H.-D. Tran, W. Xiang, and T. T. Johnson, “Verification approaches for learning-enabled autonomous cyber-physical systems,” IEEE Design & Test
2020
Later among the works it cites.
J. Fan, C. Huang, X. Chen, W. Li, and Q. Zhu, “Reachnn*: A tool for reachability analysis of neural-network controlled systems,” in International Symposium on Automated Technology for Verification and Analysis
2020
Later among the works it cites.
W. Xiang, H.-D. Tran, X. Yang, and T. T. Johnson, “Reachable set estimation for neural network control systems: A simulation-guided approach,” IEEE Transactions on Neural Networks and Learning Systems
2020
Later among the works it cites.
H. Hu, M. Fazlyab, M. Morari, and G. J. Pappas, “Reach-sdp: Reachability analysis of closed-loop systems with neural network controllers via semidefinite programming,” in 59th IEEE Conference on Decision and Control
2020
Later among the works it cites.
H. Zhang, H. Chen, C. Xiao, S. Gowal, R. Stanforth, B. Li, D. Boning, and C.-J. Hsieh, “Crown-ibp: Towards stable and efficient training of verifiably robust neural networks.” https://github.com/huanzhang12/CROWN-IBP , 2020
2020
Later among the works it cites.
S. Dathathri, K. Dvijotham, A. Kurakin, A. Raghunathan, J. Uesato, R. Bunel, S. Shankar, J. Steinhardt, I. Goodfellow, P. Liang, et al
2020
Later among the works it cites.
C. Tjandraatmadja, R. Anderson, J. Huchette, W. Ma, K. K. PATEL, and J. P. Vielma, “The convex relaxation barrier, revisited: Tightened single-neuron relaxations for neural network verification,” Advances in Neural Information Processing Systems
2020
Later among the works it cites.
B. G. Anderson, Z. Ma, J. Li, and S. Sojoudi, “Tightened convex relaxations for neural network robustness certification,” in 2020 59th IEEE Conference on Decision and Control (CDC)
2020
Later among the works it cites.
M. Everett, B. Lutjens, and J. P. How, “Certifiable robustness to adversarial state uncertainty in deep reinforcement learning,” IEEE Transactions on Neural Networks and Learning Systems
2021
Closest in time.
M. Everett, G. Habibi, and J. P. How, “Efficient reachability analysis of closed-loop systems with neural network controllers,” in IEEE International Conference on Robotics and Automation (ICRA)
2021
Closest in time.
2021
Closest in time.
S. Bansal and C. Tomlin, “DeepReach: A deep learning approach to high-dimensional reachability,” in IEEE International Conference on Robotics and Automation (ICRA)
2021
Closest in time.
S. Chen, M. Fazlyab, M. Morari, G. J. Pappas, and V. M. Preciado, “Learning lyapunov functions for hybrid systems,” in Proceedings of the 24th International Conference on Hybrid Systems: Computation and Control
2021
Closest in time.
H. Yin, P. Seiler, and M. Arcak, “Stability analysis using quadratic constraints for systems with neural network controllers,” IEEE Transactions on Automatic Control
2021
Closest in time.
J. A. Vincent and M. Schwager, “Reachable polyhedral marching (rpm): A safety verification algorithm for robotic systems with deep neural network components,” 2021
2021
Closest in time.
2021
Closest in time.
I. Ilahi, M. Usama, J. Qadir, M. U. Janjua, A. Al-Fuqaha, D. T. Huang, and D. Niyato, “Challenges and countermeasures for adversarial attacks on deep reinforcement learning,” IEEE Transactions on Artificial Intelligence
2021
Closest in time.