Fetching the paper…
Reading the bibliography…
How do neural network image classifiers respond to simpler and simpler inputs? And what do such responses reveal about the learning process? To answer these questions, we need a clear measure of input simplicity (or inversely, complexity), an optimization objective that correlates with simplification, and a framework to incorporate such objective into training and inference.
Mimic-cxr-jpg, a large publicly available database of labeled chest radiographs
Johnson, A. E., Pollard, T. J., Greenbaum, N. R., Lungren, M. P., Deng, C.-y., Peng, Y., Lu, Z., Mark, R. G., Berkowitz, S. J., and Horng, S · 1901
Earlier work this paper cites.
A mathematical theory of communication
Shannon, C. E · 1948
Earlier work this paper cites.
Pleural effusion: explanation of some typical appearances
Raasch, B., Carsky, E., Lane, E., O’Callaghan, J., and Heitzman, E · 1982
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
Hinton, G. E. and van Camp, D · 1993
Earlier work this paper cites.
A database for handwritten text recognition research
Hull, J. J · 1994
Earlier work this paper cites.
Sex differences in thoracic dimensions and configuration
Bellemare, F., Jeanneret, A., and Couture, J · 2003
Earlier work this paper cites.
80 million tiny images: A large data set for nonparametric object and scene recognition
Torralba, A., Fergus, R., and Freeman, W. T · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A · 2009
Earlier work this paper cites.
MNIST handwritten digit database
LeCun, Y. and Cortes, C · 2010
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Netzer, Y., Wang, T., Coates, A., Bissacco, A., Wu, B., and Ng, A. Y · 2011
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2015
Earlier work this paper cites.
Gradient-based hyperparameter optimization through reversible learning
Maclaurin, D., Duvenaud, D., and Adams, R · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O., Fischer, P., and Brox, T · 2015
Earlier work this paper cites.
Deep networks with stochastic depth
Huang, G., Sun, Y., Liu, Z., Sedra, D., and Weinberger, K. Q · 2016
Earlier work this paper cites.
Wide residual networks
Zagoruyko, S. and Komodakis, N · 2016
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S · 2017
Earlier work this paper cites.
SGDR: stochastic gradient descent with warm restarts
Loshchilov, I. and Hutter, F · 2017
Earlier work this paper cites.
Methods for interpreting and understanding deep neural networks
Montavon, G., Samek, W., and Müller, K.-R · 2017
Cited alongside, same era.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms, 2017
Xiao, H., Rasul, K., and Vollgraf, R · 2017
Cited alongside, same era.
Visualizing deep neural network decisions: Prediction difference analysis
Zintgraf, L. M., Cohen, T. S., Adel, T., and Welling, M · 2017
Cited alongside, same era.
Glow: Generative flow with invertible 1x1 convolutions
Kingma, D. P. and Dhariwal, P · 2018
Cited alongside, same era.
Wang, T., Zhu, J., Torralba, A., and Efros, A. A · 2018
Cited alongside, same era.
Counterfactual visual explanations
Goyal, Y., Wu, Z., Ernst, J., Batra, D., Parikh, D., and Lee, S · 2019
Estimating example difficulty using variance of gradients, 2021
Agarwal, C., D’souza, D., and Hooker, S · 2021
Later among the works it cites.
"will you find these shortcuts?" a protocol for evaluating the faithfulness of input salience methods for text classification, 2021
Bastings, J., Ebert, S., Zablotskaia, P., Sandholm, A., and Filippova, K · 2021
Later among the works it cites.
High-performance large-scale image recognition without normalization
Brock, A., De, S., Smith, S. L., and Simonyan, K · 2021
Later among the works it cites.
Diffeomorphic explanations with normalizing flows
Dombrowski, A.-K., Gerken, J. E., and Kessel, P · 2021
Later among the works it cites.
A tale of two long tails, 2021
D’souza, D., Nussbaum, Z., Agarwal, C., and Hooker, S · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A benchmark for interpretability methods in deep neural networks
Hooker, S., Erhan, D., Kindermans, P.-J., and Kim, B · 2019
Cited alongside, same era.
Pleural effusion in adults—etiology, diagnosis, and treatment
Jany, B. and Welte, T · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Loshchilov, I. and Hutter, F · 2019
Cited alongside, same era.
Shortcut learning in deep neural networks
Geirhos, R., Jacobsen, J.-H., Michaelis, C., Zemel, R., Brendel, W., Bethge, M., and Wichmann, F. A · 2020
Cited alongside, same era.
Why normalizing flows fail to detect out-of-distribution data
Kirichenko, P., Izmailov, P., and Wilson, A. G · 2020
Cited alongside, same era.
Differences between human and machine perception in medical diagnosis
Makino, T., Jastrzebski, S., Oleszkiewicz, W., Chacko, C., Ehrenpreis, R., Samreen, N., Chhor, C., Kim, E., Lee, J., Pysarenko, K., et al · 2020
Cited alongside, same era.
Dubois, Y., Bloem-Reddy, B., Ullrich, K., and Maddison, C. J · 2021
Later among the works it cites.
Improving performance of deep learning models with axiomatic attribution priors and expected gradients
Erion, G., Janizek, J. D., Sturmfels, P., Lundberg, S. M., and Lee, S.-I · 2021
Later among the works it cites.
Hierarchical vaes know what they don’t know
Havtorn, J. D., Frellsen, J., Hauberg, S., and Maaløe, L · 2021
Later among the works it cites.
ECINN: efficient counterfactuals from invertible neural networks
Hvilshøj, F., Iosifidis, A., and Assent, I · 2021
Later among the works it cites.
Evaluating the faithfulness of importance measures in nlp by recursively masking allegedly important tokens and retraining, 2021
Madsen, A., Meade, N., Adlakha, V., and Reddy, S · 2021
Later among the works it cites.
Dataset distillation with infinitely wide convolutional networks
Nguyen, T., Novak, R., Xiao, L., and Lee, J · 2021
Later among the works it cites.
Meta pseudo labels
Pham, H., Dai, Z., Xie, Q., and Le, Q. V · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al · 2021
Later among the works it cites.
Teaching with commentaries
Raghu, A., Raghu, M., Kornblith, S., Duvenaud, D., and Hinton, G. E · 2021
Later among the works it cites.
Dataset condensation with differentiable siamese augmentation
Zhao, B. and Bilen, H · 2021
Later among the works it cites.
Dataset condensation with gradient matching
Zhao, B., Mopuri, K. R., and Bilen, H · 2021
Later among the works it cites.