Fetching the paper…
Reading the bibliography…
Major winning Convolutional Neural Networks (CNNs), such as AlexNet, VGGNet, ResNet, GoogleNet, include tens to hundreds of millions of parameters, which impose considerable computation and memory overhead.
Polynomial theory of complex systems
Ivakhnenko, A. G · 1971
Earlier work this paper cites.
Neural network model for a mechanism of pattern recognition unaffected by shift in position- neocognitron
Fukushima, Kunihiko · 1979
Earlier work this paper cites.
Neocognitron: A self-organizing neural network model for a mechanism of pattern recognition unaffected by shift in position
Fukushima, Kunihiko · 1980
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Lecun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Model compression
Buciluǎ, Cristian, Caruana, Rich, and Niculescu-Mizil, Alexandru · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A. and Hinton, G · 2009
Earlier work this paper cites.
Deep, big, simple neural nets for handwritten digit recognition
Ciresan, Dan Claudiu, Meier, Ueli, Gambardella, Luca Maria, and Schmidhuber, Jürgen · 2010
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Glorot, Xavier and Bengio, Yoshua · 2010
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, Vinod and Hinton, Geoffrey E · 2010
Earlier work this paper cites.
A committee of neural networks for traffic sign classification
Cireşan, Dan, Meier, Ueli, Masci, Jonathan, and Schmidhuber, Jürgen · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Netzer, Yuval, Wang, Tao, Coates, Adam, Bissacco, Alessandro, Wu, Bo, and Ng, Andrew Y · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex, Krizhevsky, Sutskever, Ilya, and Geoffrey, E. Hinton · 2012
Earlier work this paper cites.
Multi-column deep neural networks for image classification
Ciregan, Dan, Meier, Ueli, and Schmidhuber, Jürgen · 2012
Earlier work this paper cites.
Multi-column deep neural network for traffic sign classification
CireşAn, Dan, Meier, Ueli, Masci, Jonathan, and Schmidhuber, Jürgen · 2012
Earlier work this paper cites.
Improving neural networks by preventing co-adaptation of feature detectors
Hinton, Geoffrey E, Srivastava, Nitish, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan R · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Bengio, Yoshua, Courville, Aaron, and Vincent, Pascal · 2013
Earlier work this paper cites.
Deep learning with cots hpc systems
Coates, Adam, Huval, Brody, Wang, Tao, Wu, David, Catanzaro, Bryan, and Andrew, Ng · 2013
Earlier work this paper cites.
Maxout networks
Goodfellow, Ian J, Warde-Farley, David, Mirza, Mehdi, Courville, Aaron, and Bengio, Yoshua · 2013
Earlier work this paper cites.
Building high-level features using large scale unsupervised learning
Le, Quoc V · 2013
Earlier work this paper cites.
Lin, Min, Chen, Qiang, and Yan, Shuicheng · 2013
Cited alongside, same era.
Rectifier nonlinearities improve neural network acoustic models
Maas, Andrew L, Hannun, Awni Y, and Ng, Andrew Y · 2013
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Saxe, Andrew M, McClelland, James L, and Ganguli, Surya · 2013
Cited alongside, same era.
Dropout training as adaptive regularization
Wager, Stefan, Wang, Sida, and Liang, Percy S · 2013
Cited alongside, same era.
Regularization of neural networks using dropconnect
Wan, Li, Zeiler, Matthew, Zhang, Sixin, Cun, Yann L, and Fergus, Rob · 2013
Cited alongside, same era.
Do deep nets really need to be deep?
Mishkin, Dmytro and Matas, Jiri · 2015
Later among the works it cites.
Imagenet large scale visual recognition challenge
Russakovsky, Olga, Deng, Jia, Su, Hao, Krause, Jonathan, Satheesh, Sanjeev, Ma, Sean, Huang, Zhiheng, Karpathy, Andrej, Khosla, Aditya, Bernstein, Michael, Berg, Alexander C., and Fei-Fei, Li · 2015
Later among the works it cites.
Apac: Augmented pattern classification with neural networks
Sato, Ikuro, Nishimura, Hiroki, and Yokoi, Kensuke · 2015
Later among the works it cites.
Very deep multilingual convolutional neural networks for lvcsr
Sercu, Tom, Puhrsch, Christian, Kingsbury, Brian, and LeCun, Yann · 2015
Later among the works it cites.
Scalable bayesian optimization using deep neural networks
Snoek, Jasper, Rippel, Oren, Swersky, Kevin, Kiros, Ryan, Satish, Nadathur, Sundaram, Narayanan, Patwary, Mostofa, Ali, Mostofa, and Adams, Ryan P · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ba, Jimmy and Caruana, Rich · 2014
Cited alongside, same era.
Caffe: Convolutional architecture for fast feature embedding
Jia, Yangqing, Shelhamer, Evan, Donahue, Jeff, Karayev, Sergey, Long, Jonathan, Girshick, Ross, Guadarrama, Sergio, and Darrell, Trevor · 2014
Cited alongside, same era.
Fitnets: Hints for thin deep nets
Romero, Adriana, Ballas, Nicolas, Kahou, Samira Ebrahimi, Chassang, Antoine, Gatta, Carlo, and Bengio, Yoshua · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, Karen and Zisserman, Andrew · 2014
Cited alongside, same era.
Striving for simplicity: The all convolutional net
Springenberg, Jost Tobias, Dosovitskiy, Alexey, Brox, Thomas, and Riedmiller, Martin A · 2014
Cited alongside, same era.
Deep networks with internal selective attention through feedback connections
Stollenga, Marijn F, Masci, Jonathan, Gomez, Faustino, and Schmidhuber, Jürgen · 2014
Cited alongside, same era.
Fast and accurate deep network learning by exponential linear units (elus)
Clevert, Djork-Arné, Unterthiner, Thomas, and Hochreiter, Sepp · 2015
Cited alongside, same era.
Srivastava, Rupesh Kumar, Greff, Klaus, and Schmidhuber, Jürgen · 2015
Later among the works it cites.
Going deeper with convolutions
Szegedy, Christian, Liu, Wei, Jia, Yangqing, Sermanet, Pierre, Reed, Scott, Anguelov, Dragomir, Erhan, Dumitru, Vanhoucke, Vincent, and Rabinovich, Andrew · 2015
Later among the works it cites.
Deep image: Scaling up image recognition
Wu, Ren, Yan, Shengen, Shan, Yi, Dang, Qingqing, and Sun, Gang · 2015
Later among the works it cites.
Empirical evaluation of rectified activations in convolutional network
Xu, Bing, Wang, Naiyan, Chen, Tianqi, and Li, Mu · 2015
Later among the works it cites.
Understanding neural networks through deep visualization
Yosinski, Jason, Clune, Jeff, Nguyen, Anh, Fuchs, Thomas, and Lipson, Hod · 2015
Later among the works it cites.
92.45% on cifar-10 in torch
Zagoruyko, Sergey · 2015
Later among the works it cites.
All you need is a good init
Dmytro Mishkin, Jiri Matas · 2016
Closest in time.
Deep networks with stochastic depth
Huang, Gao, Sun, Yu, Liu, Zhuang, Sedra, Daniel, and Weinberger, Kilian · 2016
Closest in time.
Squeezenet: Alexnet-level accuracy with 50x fewer parameters and¡ 0.5 mb model size
Iandola, Forrest N, Han, Song, Moskewicz, Matthew W, Ashraf, Khalid, Dally, William J, and Keutzer, Kurt · 2016
Closest in time.
Generalizing pooling functions in convolutional neural networks: Mixed, gated, and tree
Lee, Chen-Yu, Gallagher, Patrick W, and Tu, Zhuowen · 2016
Closest in time.
Inception-v4, inception-resnet and the impact of residual connections on learning
Szegedy, Christian, Ioffe, Sergey, Vanhoucke, Vincent, and Alemi, Alex · 2016
Closest in time.
Zagoruyko, Sergey and Komodakis, Nikos · 2016
Closest in time.
Pytorch: An imperative style, high-performance deep learning library, 2019
Paszke, Adam, Gross, Sam, Massa, Francisco, Lerer, Adam, Bradbury, James, Chanan, Gregory, Killeen, Trevor, Lin, Zeming, Gimelshein, Natalia, Antiga, Luca, Desmaison, Alban, Köpf, Andreas, Yang, Edward, DeVito, Zach, Raison, Martin, Tejani, Alykhan, Chilamkurthy, Sasank, Steiner, Benoit, Fang, Lu, Bai, Junjie, and Chintala, Soumith · 2019
Closest in time.
Pytorch image models
Wightman, Ross · 2019
Closest in time.