Fetching the paper…
Reading the bibliography…
In recent years, state-of-the-art methods in computer vision have utilized increasingly deep convolutional neural network architectures (CNNs), with some of the most successful models employing hundreds or even thousands of layers.
A matrix approach to discrete wavelets
Kautsky, J. and Turcajová, R · 1994
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Natural language processing (almost) from scratch
Collobert, R., Weston, J., Bottou, L., Karlen, M., Kavukcuoglu, K., and Kuksa, P · 2011
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Hinton, G., Deng, L., Yu, D., Dahl, G. E., Mohamed, A.-r., Jaitly, N., Senior, A., Vanhoucke, V., Nguyen, P., Sainath, T. N., et al · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Earlier work this paper cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Saxe, A. M., McClelland, J. L., and Ganguli, S · 2013
Earlier work this paper cites.
A convolutional neural network for modelling sentences
Kalchbrenner, N., Grefenstette, E., and Blunsom, P · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Kim, Y · 2014
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S. and Szegedy, C · 2015
Cited alongside, same era.
Mishkin, D. and Matas, J · 2015
Cited alongside, same era.
Toward deeper understanding of neural networks: The power of initialization and a dual view on expressivity
Daniely, A., Frostig, R., and Singer, Y · 2016
Cited alongside, same era.
Exponential expressivity in deep neural networks through transient chaos
Poole, B., Lahiri, S., Raghu, M., Sohl-Dickstein, J., and Ganguli, S · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., van den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., Dieleman, S., Grewe, D., Nham, J., Kalchbrenner, N., Sutskever, I., Lillicrap, T., Leach, M., Kavukcuoglu, K., Graepel, T., and Hassabis, D · 2016
Cited alongside, same era.
Mastering the game of go without human knowledge
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., et al · 2017
Later among the works it cites.
Mean field residual networks: On the edge of chaos
Yang, G. and Schoenholz, S · 2017
Later among the works it cites.
How to start training: The effect of initialization and architecture
Hanin, B. and Rolnick, D · 2018
Closest in time.
On the selection of initialization and activation function for deep neural networks
Hayou, S., Doucet, A., and Rousseau, J · 2018
Closest in time.
Universal Statistics of Fisher Information in Deep Neural Networks: Mean Field Approach
Karakida, R., Akaho, S., and Amari, S.-i · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Resurrecting the sigmoid in deep learning through dynamical isometry: theory and practice
Pennington, J., Schoenholz, S., and Ganguli, S · 2017
Cited alongside, same era.
Deep Information Propagation
Schoenholz, S. S., Gilmer, J., Ganguli, S., and Sohl-Dickstein, J · 2017
Cited alongside, same era.
A correspondence between random neural networks and statistical field theory
Schoenholz, S. S., Pennington, J., and Sohl-Dickstein, J · 2017
Cited alongside, same era.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J
Cited in the paper.
Identity mappings in deep residual networks
He, K., Zhang, X., Ren, S., and Sun, J
Cited in the paper.
The emergence of spectral universality in deep networks
Pennington, J., Schoenholz, S. S., and Ganguli, S · 2018
Closest in time.
Deep mean field theory: Layerwise variance and width variation as methods to control gradient explosion, 2018
Yang, G. and Schoenholz, S. S · 2018
Closest in time.