Fetching the paper…
Reading the bibliography…
One major challenge in training Deep Neural Networks is preventing overfitting.
On the stability of inverse problems
Tikhonov, Andrey Nikolayevich · 1943
Earlier work this paper cites.
Possible principles underlying the transformations of sensory messages
Barlow, Horace B · 1961
Earlier work this paper cites.
G-maximization: An unsupervised learning procedure for discovering regularities
Pearlmutter, Barak A and Hinton, Geoffrey · 1986
Earlier work this paper cites.
Self-organization in a perceptual network
Linsker, Ralph · 1988
Earlier work this paper cites.
Neural network ensembles
Hansen, Lars Kai and Salamon, Peter · 1990
Earlier work this paper cites.
Learning factorial codes by predictability minimization
Schmidhuber, Jürgen · 1992
Earlier work this paper cites.
When networks disagree: Ensemble methods for hybrid neural networks
Perrone, Michael P. and Cooper, Leaon N · 1993
Earlier work this paper cites.
Comparison of learning algorithms for handwritten digit recognition
LeCun, Yann, Jackel, LD, Bottou, L, Brunot, A, Cortes, C, Denker, JS, Drucker, H, Guyon, I, Muller, UA, Sackinger, E, et al · 1995
Earlier work this paper cites.
Bagging predictors
Breiman, Leo · 1996
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Tibshirani, Robert · 1996
Earlier work this paper cites.
Slow, decorrelated features for pretraining complex cell-like networks
Bengio, Yoshua and Bergstra, James S · 2009
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Cited alongside, same era.
Learning multiple layers of features from tiny images
Krizhevsky, Alex and Hinton, Geoffrey · 2009
Cited alongside, same era.
Understanding the difficulty of training deep feedforward neural networks
Glorot, Xavier and Bengio, Yoshua · 2010
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoff · 2012
Cited alongside, same era.
Deep canonical correlation analysis
Andrew, Galen, Arora, Raman, Bilmes, Jeff, and Livescu, Karen · 2013
Cited alongside, same era.
Maxout networks
Goodfellow, Ian J, Warde-Farley, David, Mirza, Mehdi, Courville, Aaron, and Bengio, Yoshua · 2013
Cited alongside, same era.
Rich feature hierarchies for accurate object detection and semantic segmentation
Girshick, Ross, Donahue, Jeff, Darrell, Trevor, and Malik, Jagannath · 2014
Later among the works it cites.
Dropout: A simple way to prevent neural networks from overfitting
Srivastava, Nitish, Hinton, Geoffrey, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan · 2014
Later among the works it cites.
How transferable are features in deep neural networks?
Yosinski, Jason, Clune, Jeff, Bengio, Yoshua, and Lipson, Hod · 2014
Later among the works it cites.
Learning deep features for scene recognition using places database
Zhou, B., Lapedriza, A., Xiao, J., Torralba, A., and Oliva, A · 2014
Later among the works it cites.
Vqa: Visual question answering
Antol, Stanislaw, Agrawal, Aishwarya, Lu, Jiasen, Mitchell, Margaret, Batra, Dhruv, Zitnick, C. Lawrence, and Parikh, Devi · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Caffe: An open source convolutional architecture for fast feature embedding
Jia, Yangqing · 2013
Cited alongside, same era.
Regularization of neural networks using dropconnect
Wan, Li, Zeiler, Matthew, Zhang, Sixin, Cun, Yann L, and Fergus, Rob · 2013
Cited alongside, same era.
Discovering hidden factors of variation in deep networks
Cheung, Brian, Livezey, Jesse A., Bansal, Arjun K., and Olshausen, Bruno A · 2014
Cited alongside, same era.
Decaf: A deep convolutional activation feature for generic visual recognition
Donahue, Jeff, Jia, Yangqing, Vinyals, Oriol, Hoffman, Judy, Zhang, Ning, Tzeng, Eric, and Darrell, Trevor · 2014
Cited alongside, same era.
Network in network
Lin, Min, Chen, Qiang, and Yan, Shuicheng
Cited in the paper.
Microsoft COCO: Common objects in context, 2014b
Lin, Tsung-Yi, Maire, Michael, Belongie, Serge, Hays, James, Perona, Pietro, Ramanan, Deva, Doll�r, Piotr, and Zitnick, C. Lawrence
Cited in the paper.
Chandar, Sarath, Khapra, Mitesh M, Larochelle, Hugo, and Ravindran, Balaraman · 2015
Closest in time.
Mind’s eye: A recurrent visual representation for image caption generation
Chen, Xinlei and Zitnick, C Lawrence · 2015
Closest in time.
Fast r-cnn
Girshick, Ross · 2015
Closest in time.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, Sergey and Szegedy, Christian · 2015
Closest in time.
Show and tell: A neural image caption generator
Vinyals, Oriol, Toshev, Alexander, Bengio, Samy, and Erhan, Dumitru · 2015
Closest in time.