Fetching the paper…
Reading the bibliography…
Detecting anomalous inputs, such as adversarial and out-of-distribution (OOD) inputs, is critical for classifiers (including deep neural networks or DNNs) deployed in real-world applications.
Statistical methods for research workers
Fisher, R. A · 1992
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Anomalous instance detection in deep learning: A survey
Bulusu, S., Kailkhura, B., Li, B., Varshney, P. K., and Song, D · 2003
Earlier work this paper cites.
Neighborhood preserving embedding
He, X., Cai, D., Yan, S., and Zhang, H.-J · 2005
Earlier work this paper cites.
Can machine learning be secure?
Barreno, M., Nelson, B., Sears, R., Joseph, A. D., and Tygar, J. D · 2006
Earlier work this paper cites.
The relationship between Precision-Recall and ROC curves
Davis, J. and Goadrich, M · 2006
Earlier work this paper cites.
An introduction to ROC analysis
Fawcett, T · 2006
Earlier work this paper cites.
Multiple testing procedures with applications to genomics
Dudoit, S. and Van Der Laan, M. J · 2007
Earlier work this paper cites.
Growing a multi-class classifier with a reject option
Tax, D. M. J. and Duin, R. P. W · 2008
Earlier work this paper cites.
Anomaly detection: A survey
Chandola, V., Banerjee, A., and Kumar, V · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A., Hinton, G., et al · 2009
Earlier work this paper cites.
Anomaly detection with score functions based on nearest neighbor graphs
Zhao, M. and Saligrama, V · 2009
Earlier work this paper cites.
NotMNIST dataset
Bulatov, Y · 2011
Earlier work this paper cites.
Efficient k-nearest neighbor graph construction for generic similarity measures
Dong, W., Moses, C., and Li, K · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Netzer, Y., Wang, T., Coates, A., Bissacco, A., Wu, B., and Ng, A. Y · 2011
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
Pedregosa, F., Varoquaux, G., Gramfort, A., Michel, V., Thirion, B., Grisel, O., Blondel, M., Prettenhofer, P., Weiss, R., Dubourg, V., et al · 2011
Earlier work this paper cites.
Bayesian reasoning and machine learning
Barber, D · 2012
Earlier work this paper cites.
Poisoning attacks against support vector machines
Biggio, B., Nelson, B., and Laskov, P · 2012
Earlier work this paper cites.
New statistic in p-value estimation for anomaly detection
Qian, J. and Saligrama, V · 2012
Earlier work this paper cites.
Goodness-of-fit statistics for discrete multivariate data
Read, T. R. and Cressie, N. A · 2012
Earlier work this paper cites.
Intriguing properties of neural networks
Szegedy, C., Zaremba, W., Sutskever, I., Bruna, J., Erhan, D., Goodfellow, I. J., and Fergus, R · 2014
Earlier work this paper cites.
Estimating local intrinsic dimensionality
Amsaleg, L., Chelly, O., Furon, T., Girard, S., Houle, M. E., Kawarabayashi, K.-i., and Nett, M · 2015
Earlier work this paper cites.
Precision-recall-gain curves: PR analysis done right
Flach, P. and Kull, M · 2015
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Goodfellow, I. J., Shlens, J., and Szegedy, C · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, K., Zhang, X., Ren, S., and Sun, J · 2015
Cited alongside, same era.
Deep neural networks are easily fooled: High confidence predictions for unrecognizable images
Nguyen, A. M., Yosinski, J., and Clune, J · 2015
Cited alongside, same era.
Robustness of classifiers: from adversarial to random noise
Fawzi, A., Moosavi-Dezfooli, S., and Frossard, P · 2016
Cited alongside, same era.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Cited alongside, same era.
Deepfool: A simple and accurate method to fool deep neural networks
Moosavi-Dezfooli, S., Fawzi, A., and Frossard, P · 2016
Obfuscated gradients give a false sense of security: Circumventing defenses to adversarial examples
Athalye, A., Carlini, N., and Wagner, D. A · 2018
Later among the works it cites.
Wild patterns: Ten years after the rise of adversarial machine learning
Biggio, B. and Roli, F · 2018
Later among the works it cites.
Analysis of classifiers’ robustness to adversarial perturbations
Fawzi, A., Fawzi, O., and Frossard, P · 2018
Later among the works it cites.
To trust or not to trust a classifier
Jiang, H., Kim, B., Guan, M. Y., and Gupta, M. R · 2018
Later among the works it cites.
A simple unified framework for detecting out-of-distribution samples and adversarial attacks
Lee, K., Lee, K., Lee, H., and Shin, J · 2018
Later among the works it cites.
Characterizing adversarial subspaces using local intrinsic dimensionality
Ma, X., Li, B., Wang, Y., Erfani, S. M., Wijewickrema, S. N. R., Schoenebeck, G., Song, D., Houle, M. E., and Bailey, J · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The limitations of deep learning in adversarial settings
Papernot, N., McDaniel, P. D., Jha, S., Fredrikson, M., Celik, Z. B., and Swami, A · 2016
Cited alongside, same era.
Learning minimum volume sets and anomaly detectors from KNN graphs
Root, J., Saligrama, V., and Qian, J · 2016
Cited alongside, same era.
An overview of gradient descent optimization algorithms
Ruder, S · 2016
Cited alongside, same era.
Towards evaluating the robustness of neural networks
Carlini, N. and Wagner, D. A · 2017
Cited alongside, same era.
Detecting adversarial samples from artifacts
Feinman, R., Curtin, R. R., Shintre, S., and Gardner, A. B · 2017
Cited alongside, same era.
A baseline for detecting misclassified and out-of-distribution examples in neural networks
Hendrycks, D. and Gimpel, K · 2017
Cited alongside, same era.
Later among the works it cites.
Towards deep learning models resistant to adversarial attacks
Madry, A., Makelov, A., Schmidt, L., Tsipras, D., and Vladu, A · 2018
Later among the works it cites.
Deep k-nearest neighbors: Towards confident, interpretable and robust deep learning
Papernot, N. and McDaniel, P. D · 2018
Later among the works it cites.
Feature squeezing: Detecting adversarial examples in deep neural networks
Xu, W., Evans, D., and Qi, Y · 2018
Later among the works it cites.
Adversarial examples: Attacks and defenses for deep learning
Yuan, X., He, P., Zhu, Q., and Li, X · 2018
Later among the works it cites.
Robust detection of adversarial attacks by modeling the intrinsic properties of deep neural networks
Zheng, Z. and Hong, P · 2018
Later among the works it cites.
Why relu networks yield high-confidence predictions far away from the training data and how to mitigate the problem
Hein, M., Andriushchenko, M., and Bitterwolf, J · 2019
Later among the works it cites.
Attribution-based confidence metric for deep neural networks
Jha, S., Raj, S., Fernandes, S. L., Jha, S. K., Jha, S., Jalaian, B., Verma, G., and Swami, A · 2019
Later among the works it cites.
When not to classify: Anomaly detection of attacks (ADA) on DNN classifiers at test time
Miller, D. J., Wang, Y., and Kesidis, G · 2019
Later among the works it cites.
The odds are odd: A statistical test for detecting adversarial examples
Roth, K., Kilcher, Y., and Hofmann, T · 2019
Later among the works it cites.
The harmonic mean p-value for combining dependent tests
Wilson, D. J · 2019
Later among the works it cites.
Detecting semantic anomalies
Ahmed, F. and Courville, A. C · 2020
Closest in time.
Adversarial learning targeting deep neural network classification: A comprehensive review of defenses against attacks
Miller, D. J., Xiang, Z., and Kesidis, G · 2020
Closest in time.
Detecting out-of-distribution examples with gram matrices
Sastry, C. S. and Oore, S · 2020
Closest in time.
Minimum-norm adversarial examples on KNN and KNN-based models
Sitawarin, C. and Wagner, D. A · 2020
Closest in time.
On adaptive attacks to adversarial example defenses
Tramèr, F., Carlini, N., Brendel, W., and Madry, A · 2020
Closest in time.
ML-LOO: Detecting adversarial examples with feature attribution
Yang, P., Chen, J., Hsieh, C., Wang, J., and Jordan, M. I · 2020
Closest in time.