Fetching the paper…
Reading the bibliography…
Unsupervised pre-training was a critical technique for training deep neural networks years ago.
A fast learning algorithm for deep belief nets
G. E. Hinton, S. Osindero, and Y. W. Teh · 2006
Earlier work this paper cites.
Unsupervised learning of invariant feature hierarchies with applications to object recognition
M. Ranzato, F. J. Huang, Y. Boureau, and Y. LeCun · 2007
Earlier work this paper cites.
Learning deep architectures for AI
Y. Bengio · 2009
Earlier work this paper cites.
Why does unsupervised pre-training help deep learning?
D. Erhan, Y. Bengio, A. C. Courville, P. Manzagol, P. Vincent, and S. Bengio · 2010
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
X. Glorot and Y. Bengio · 2010
Earlier work this paper cites.
Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P. Manzagol · 2010
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Discriminative unsupervised feature learning with convolutional neural networks
A. Dosovitskiy, J. T. Springenberg, M. A. Riedmiller, and T. Brox · 2014
Earlier work this paper cites.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. S. Corrado, A. Davis, J. Dean, M. Devin, S. Ghemawat, I. Goodfellow, A. Harp, G. Irving, M. Isard, Y. Jia, R. Jozefowicz, L. Kaiser, M. Kudlur, J. Levenberg, D. Mané, R. Monga, S. Moore, D. Murray, C. Olah, M. Schuster, J. Shlens, B. Steiner, I. Sutskever, K. Talwar, P. Tucker, V. Vanhoucke, V. Vasudevan, F. Viégas, O. Vinyals, P. Warden, M. Wattenberg, M. Wicke, Y. Yu, and X. Zheng · 2015
Earlier work this paper cites.
Segnet: A deep convolutional encoder-decoder architecture for image segmentation
V. Badrinarayanan, A. Kendall, and R. Cipolla · 2015
Earlier work this paper cites.
Unsupervised visual representation learning by context prediction
C. Doersch, A. Gupta, and A. A. Efros · 2015
Earlier work this paper cites.
The pascal visual object classes challenge: A retrospective
M. Everingham, S. M. A. Eslami, L. J. V. Gool, C. K. I. Williams, J. M. Winn, and A. Zisserman · 2015
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
Learning deconvolution network for semantic segmentation
H. Noh, S. Hong, and B. Han · 2015
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
A. Radford, L. Metz, and S. Chintala · 2015
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
From neural PCA to deep unsupervised learning
H. Valpola · 2015
Later among the works it cites.
Unsupervised learning of visual representations using videos
X. Wang and A. Gupta · 2015
Later among the works it cites.
Deep representation learning with target coding
S. Yang, P. Luo, C. C. Loy, K. W. Shum, and X. Tang · 2015
Later among the works it cites.
Mxnet: A flexible and efficient machine learning library for heterogeneous distributed systems
T. Chen, M. Li, Y. Li, M. Lin, N. Wang, M. Wang, T. Xiao, B. Xu, C. Zhang, and Z. Zhang · 2016
Closest in time.
Fast and accurate deep network learning by exponential linear units (elus)
D. Clevert, T. Unterthiner, and S. Hochreiter · 2016
Closest in time.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Radford, L. Metz, and S. Chintala · 2015
Cited alongside, same era.
Semi-supervised learning with ladder networks
A. Rasmus, M. Berglund, M. Honkala, H. Valpola, and T. Raiko · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Cited alongside, same era.
Scalable bayesian optimization using deep neural networks
J. Snoek, O. Rippel, K. Swersky, R. Kiros, N. Satish, N. Sundaram, M. M. A. Patwary, Prabhat, and R. P. Adams · 2015
Cited alongside, same era.
Unsupervised and semi-supervised learning with categorical generative adversarial networks
J. T. Springenberg · 2015
Cited alongside, same era.
Striving for simplicity: The all convolutional net
J. T. Springenberg, A. Dosovitskiy, T. Brox, and M. A. Riedmiller · 2015
Cited alongside, same era.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. E. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Cited alongside, same era.
Identity mappings in deep residual networks
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Closest in time.
Unsupervised learning of discriminative attributes and visual representations
C. Huang, C. Change Loy, and X. Tang · 2016
Closest in time.
Image restoration using convolutional auto-encoders with symmetric skip connections
X. Mao, C. Shen, and Y. Yang · 2016
Closest in time.
Context encoders: Feature learning by inpainting
D. Pathak, P. Krähenbühl, J. Donahue, T. Darrell, and A. A. Efros · 2016
Closest in time.
Improved techniques for training gans
T. Salimans, I. J. Goodfellow, W. Zaremba, V. Cheung, A. Radford, and X. Chen · 2016
Closest in time.
S. Zagoruyko and N. Komodakis · 2016
Closest in time.