Fetching the paper…
Reading the bibliography…
Our ability to know when to trust the decisions made by machine learning systems has not kept up with the staggering improvements in their performance, limiting their applicability in high-stakes domains.
Elements of information theory
Thomas M Cover · 1999
Earlier work this paper cites.
Numerical optimization
Jorge Nocedal and Stephen Wright · 2006
Earlier work this paper cites.
Theory of games and economic behavior
John Von Neumann and Oskar Morgenstern · 2007
Earlier work this paper cites.
Probabilistic proof systems: A primer
Oded Goldreich · 2008
Earlier work this paper cites.
Computational complexity: a modern approach
Sanjeev Arora and Boaz Barak · 2009
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Andrew L Maas, Awni Y Hannun, and Andrew Y Ng · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Ian J. Goodfellow, Jonathon Shlens, and Christian Szegedy · 2015
Earlier work this paper cites.
Spatial transformer networks
Max Jaderberg, Karen Simonyan, Andrew Zisserman, et al · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton · 2016
Earlier work this paper cites.
Rationalizing neural predictions
Tao Lei, Regina Barzilay, and Tommi S. Jaakkola · 2016
Earlier work this paper cites.
Reluplex: An efficient SMT solver for verifying deep neural networks
Guy Katz, Clark Barrett, David L Dill, Kyle Julian, and Mykel J Kochenderfer · 2017
Earlier work this paper cites.
The concrete distribution: A continuous relaxation of discrete random variables
C Maddison, A Mnih, and Y Teh · 2017
Cited alongside, same era.
Training verified learners with learned verifiers
Krishnamurthy Dvijotham, Sven Gowal, Robert Stanforth, Relja Arandjelovic, Brendan O’Donoghue, Jonathan Uesato, and Pushmeet Kohli · 2018
Cited alongside, same era.
Geoffrey Irving, Paul Christiano, and Dario Amodei · 2018
Cited alongside, same era.
Differentiable abstract interpretation for provably robust neural networks
Matthew Mirman, Timon Gehr, and Martin Vechev · 2018
Cited alongside, same era.
Certified defenses against adversarial examples
Aditi Raghunathan, Jacob Steinhardt, and Percy Liang · 2018
Cited alongside, same era.
Interactive proofs (part i), 2019
Justin Thaler · 2019
Later among the works it cites.
Spatial broadcast decoder: A simple architecture for learning disentangled representations in vaes
Nicholas Watters, Loic Matthey, Christopher P Burgess, and Alexander Lerchner · 2019
Later among the works it cites.
Rethinking cooperative rationalization: Introspective extraction and complement control
Mo Yu, Shiyu Chang, Yang Zhang, and Tommi Jaakkola · 2019
Later among the works it cites.
Writeup: Progress on AI safety via. debate
Beth Barnes and Paul Christiano · 2020
Later among the works it cites.
Implicit learning dynamics in Stackelberg games: Equilibria characterization, convergence analysis, and empirical study
Tanner Fiez, Benjamin Chasnov, and Lillian Ratliff · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Uncovering surprising behaviors in reinforcement learning via worst-case analysis
Avraham Ruderman, Richard Everett, Bristy Sikder, Hubert Soyer, Jonathan Uesato, Ananya Kumar, Charlie Beattie, and Pushmeet Kohli · 2018
Cited alongside, same era.
Fast and effective robustness certification
Gagandeep Singh, Timon Gehr, Matthew Mirman, Markus Püschel, and Martin Vechev · 2018
Cited alongside, same era.
Rigorous agent evaluation: An adversarial approach to uncover catastrophic failures
Jonathan Uesato, Ananya Kumar, Csaba Szepesvari, Tom Erez, Avraham Ruderman, Keith Anderson, Krishnamurthy Dj Dvijotham, Nicolas Heess, and Pushmeet Kohli · 2018
Cited alongside, same era.
Scaling provable adversarial defenses
Eric Wong, Frank R Schmidt, Jan Hendrik Metzen, and J Zico Kolter · 2018
Cited alongside, same era.
A game theoretic approach to class-wise selective rationalization
Shiyu Chang, Yang Zhang, Mo Yu, and Tommi Jaakkola · 2019
Cited alongside, same era.
Towards robust and verified ai: Specification testing, robust training, and formal verification. DeepMind. Medium, 2019
P Kohli, S Gowal, K Dvijotham, and J Uesato · 2019
Cited alongside, same era.
When does label smoothing help?
Rafael Müller, Simon Kornblith, and Geoffrey E Hinton · 2019
Cited alongside, same era.
Marcel Hildebrandt, Jorge Andres Quintero Serna, Yunpu Ma, Martin Ringsquandl, Mitchell Joblin, and Volker Tresp · 2020
Later among the works it cites.
“flood the zone with shit”: How misinformation overwhelmed our democracy
Sean Illing · 2020
Later among the works it cites.
Prevalence of neural collapse during the terminal phase of deep learning training
Vardan Papyan, XY Han, and David L Donoho · 2020
Later among the works it cites.
On solving minimax optimization locally: A follow-the-ridge approach
Yuanhao Wang*, Guodong Zhang*, and Jimmy Ba · 2020
Later among the works it cites.
Deep verifier networks: Verification of deep discriminative models with deep generative models
Tong Che, Xiaofeng Liu, Site Li, Yubin Ge, Ruixiang Zhang, Caiming Xiong, and Yoshua Bengio · 2021
Closest in time.
Interactive proofs for verifying machine learning
Shafi Goldwasser, Guy N Rothblum, Jonathan Shafer, and Amir Yehudayoff · 2021
Closest in time.
Guodong Zhang, Yuanhao Wang, Laurent Lessard, and Roger Grosse · 2021
Closest in time.