Fetching the paper…
Reading the bibliography…
A remarkable characteristic of overparameterized deep neural networks (DNNs) is that their accuracy does not degrade when the network's width is increased.
Implicit regularization in over-parameterized neural networks
Kubo, M.; Banno, R.; Manabe, H.; and Minoji, M. 2019 · 1903
Earlier work this paper cites.
Ma, X.; Yuan, G.; Lin, S.; Li, Z.; Sun, H.; and Wang, Y. 2019 · 1905
Earlier work this paper cites.
Kernel and Deep Regimes in Overparametrized Models
Woodworth, B.; Gunasekar, S.; Lee, J.; Soudry, D.; and Srebro, N. 2019 · 1906
Earlier work this paper cites.
The generalization error of random features regression: Precise asymptotics and double descent curve
Mei, S.; and Montanari, A. 2019 · 1908
Earlier work this paper cites.
Deep double descent: Where bigger models and more data hurt
Nakkiran, P.; Kaplun, G.; Bansal, Y.; Yang, T.; Barak, B.; and Sutskever, I. 2019 · 1912
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, N.; Hinton, G. E.; Krizhevsky, A.; Sutskever, I.; and Salakhutdinov, R. 2014 · 1958
Earlier work this paper cites.
Randomness, redundancy and repair: roles and relevance to biological aging
Strehler, B. L.; and Freeman, M. R. 1980 · 1980
Earlier work this paper cites.
An hypothesis about redundancy and reliability in the brains of higher species: Analogies with genes, internal organs, and engineering systems
Glassman, R. B. 1987 · 1987
Earlier work this paper cites.
Double Trouble in Double Descent: Bias and Variance (s) in the Lazy Regime
d’Ascoli, S.; Refinetti, M.; Biroli, G.; and Krzakala, F. 2020 · 2003
Earlier work this paper cites.
Compositional Explanations of Neurons
Mu, J.; and Andreas, J. 2020 · 2006
Earlier work this paper cites.
On the Number of Linear Regions of Convolutional Neural Networks
Xiong, H.; Huang, L.; Yu, M.; Liu, L.; Zhu, F.; and Shao, L. 2020 · 2006
Earlier work this paper cites.
Robust artificial neural network architectures
Schuster, A. 2008 · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A.; Hinton, G.; et al. 2009 · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Glorot, X.; and Bengio, Y. 2010 · 2010
Earlier work this paper cites.
Efficient backprop
LeCun, Y. A.; Bottou, L.; Orr, G. B.; and Müller, K.-R. 2012 · 2012
Earlier work this paper cites.
Low-rank matrix factorization for deep neural network training with high-dimensional output targets
Sainath, T. N.; Kingsbury, B.; Sindhwani, V.; Arisoy, E.; and Ramabhadran, B. 2013 · 2013
Earlier work this paper cites.
Exploiting linear structure within convolutional networks for efficient evaluation
Denton, E. L.; Zaremba, W.; Bruna, J.; LeCun, Y.; and Fergus, R. 2014 · 2014
Earlier work this paper cites.
State-building: governance and world order in the 21st century
Fukuyama, F. 2014 · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Zeiler, M. D.; and Fergus, R. 2014 · 2014
Earlier work this paper cites.
Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding
Han, S.; Mao, H.; and Dally, W. J. 2015 · 2015
Earlier work this paper cites.
Learning both weights and connections for efficient neural network
Han, S.; Pool, J.; Tran, J.; and Dally, W. 2015 · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2015 · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Hinton, G.; Vinyals, O.; and Dean, J. 2015 · 2015
Cited alongside, same era.
ImageNet Large Scale Visual Recognition Challenge
Russakovsky, O.; Deng, J.; Su, H.; Krause, J.; Satheesh, S.; Ma, S.; Huang, Z.; Karpathy, A.; Khosla, A.; Bernstein, M.; Berg, A. C.; and Fei-Fei, L. 2015 · 2015
Cited alongside, same era.
Data-free parameter pruning for deep neural networks
Srinivas, S.; and Babu, R. V. 2015 · 2015
Cited alongside, same era.
Stronger generalization bounds for deep nets via a compression approach
Arora, S.; Ge, R.; Neyshabur, B.; and Zhang, Y. 2018 · 2018
Later among the works it cites.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A.; Gabriel, F.; and Hongler, C. 2018 · 2018
Later among the works it cites.
Measuring the intrinsic dimension of objective landscapes
Li, C.; Farkhoor, H.; Liu, R.; and Yosinski, J. 2018 · 2018
Later among the works it cites.
Methods for interpreting and understanding deep neural networks
Montavon, G.; Samek, W.; and Müller, K.-R. 2018 · 2018
Later among the works it cites.
Sensitivity and Generalization in Neural Networks: an Empirical Study
Novak, R.; Bahri, Y.; Abolafia, D. A.; Pennington, J.; and Sohl-Dickstein, J. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Aerts, H.; Fias, W.; Caeyenberghs, K.; and Marinazzo, D. 2016 · 2016
Cited alongside, same era.
Artificial neural networks as models of robustness in development and regeneration: stability of memory during morphological remodeling
Hammelman, J.; Lobo, D.; and Levin, M. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Cited alongside, same era.
Network trimming: A data-driven neuron pruning approach towards efficient deep architectures
Hu, H.; Peng, R.; Tai, Y.-W.; and Tang, C.-K. 2016 · 2016
Cited alongside, same era.
Rethinking the inception architecture for computer vision
Szegedy, C.; Vanhoucke, V.; Ioffe, S.; Shlens, J.; and Wojna, Z. 2016 · 2016
Cited alongside, same era.
Using goal-driven deep learning models to understand sensory cortex
Yamins, D. L.; and DiCarlo, J. J. 2016 · 2016
Cited alongside, same era.
A closer look at memorization in deep networks
Arpit, D.; Jastrzebski, S.; Ballas, N.; Krueger, D.; Bengio, E.; Kanwal, M. S.; Maharaj, T.; Fischer, A.; Courville, A.; Bengio, Y.; et al. 2017 · 2017
Cited alongside, same era.
Olah, C.; Satyanarayan, A.; Johnson, I.; Carter, S.; Schubert, L.; Ye, K.; and Mordvintsev, A. 2018 · 2018
Later among the works it cites.
Clip-q: Deep network compression learning by in-parallel pruning-quantization
Tung, F.; and Mori, G. 2018 · 2018
Later among the works it cites.
Revisiting the importance of individual units in cnns via ablation
Zhou, B.; Sun, Y.; Bau, D.; and Torralba, A. 2018 · 2018
Later among the works it cites.
Intrinsic dimension of data representations in deep neural networks
Ansuini, A.; Laio, A.; Macke, J. H.; and Zoccolan, D. 2019 · 2019
Closest in time.
On Lazy Training in Differentiable Programming
Chizat, L.; Oyallon, E.; and Bach, F. 2019 · 2019
Closest in time.
Random deep neural networks are biased towards simple functions
De Palma, G.; Kiani, B.; and Lloyd, S. 2019 · 2019
Closest in time.
The lottery ticket hypothesis: Finding sparse, trainable neural networks
Frankle, J.; and Carbin, M. 2019 · 2019
Closest in time.
Implicit regularization of discrete gradient dynamics in deep linear neural networks
Gidel, G.; Bach, F.; and Lacoste-Julien, S. 2019 · 2019
Closest in time.
Deep relu networks have surprisingly few activation patterns
Hanin, B.; and Rolnick, D. 2019 · 2019
Closest in time.
The role of over-parametrization in generalization of neural networks
Neyshabur, B.; Li, Z.; Bhojanapalli, S.; LeCun, Y.; and Srebro, N. 2019 · 2019
Closest in time.
Non-vacuous generalization bounds at the imagenet scale: a PAC-bayesian compression approach
Zhou, W.; Veitch, V.; Austern, M.; Adams, R. P.; and Orbanz, P. 2019 · 2019
Closest in time.
Understanding the role of individual units in a deep neural network
Bau, D.; Zhu, J.-Y.; Strobelt, H.; Lapedriza, A.; Zhou, B.; and Torralba, A. 2020 · 2020
Closest in time.
A comprehensive survey on model compression and acceleration
Choudhary, T.; Mishra, V.; Goswami, A.; and Sarangapani, J. 2020 · 2020
Closest in time.
Compression based bound for non-compressed network: unified generalization error analysis of large compressible deep neural network
Suzuki, T.; Abe, H.; and Nishimura, T. 2020 · 2020
Closest in time.