Fetching the paper…
Reading the bibliography…
We explore a recently proposed Variational Dropout technique that provided an elegant Bayesian interpretation to Gaussian Dropout.
Bayesian interpolation
MacKay, David JC · 1992
Earlier work this paper cites.
Bayesian nonlinear modeling for the prediction competition
MacKay, David JC et al · 1994
Earlier work this paper cites.
Bayesian learning for neural networks , volume 118
Neal, Radford M · 1996
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Yann, Bottou, Léon, Bengio, Yoshua, and Haffner, Patrick · 1998
Earlier work this paper cites.
Sparse bayesian learning and the relevance vector machine
Tipping, Michael E · 2001
Earlier work this paper cites.
Automatic relevance determination for least squares support vector machine regression
Van Gestel, Tony, Suykens, JAK, De Moor, Bart, and Vandewalle, Joos · 2001
Earlier work this paper cites.
Learning multiple layers of features from tiny images, 2009
Krizhevsky, Alex and Hinton, Geoffrey · 2009
Earlier work this paper cites.
Theano: A cpu and gpu math compiler in python
Bergstra, James, Breuleux, Olivier, Bastien, Frédéric, Lamblin, Pascal, Pascanu, Razvan, Desjardins, Guillaume, Turian, Joseph, Warde-Farley, David, and Bengio, Yoshua · 2010
Earlier work this paper cites.
On over-fitting in model selection and subsequent selection bias in performance evaluation
Cawley, Nicola L. C. Talbot · 2010
Earlier work this paper cites.
Improving neural networks by preventing co-adaptation of feature detectors
Hinton, Geoffrey E, Srivastava, Nitish, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan R · 2012
Earlier work this paper cites.
Gaussian kullback-leibler approximate inference
Challis, E and Barber, D · 2013
Earlier work this paper cites.
Stochastic variational inference
Hoffman, Matthew D, Blei, David M, Wang, Chong, and Paisley, John William · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, Diederik P and Welling, Max · 2013
Earlier work this paper cites.
Regularization of neural networks using dropconnect
Wan, Li, Zeiler, Matthew, Zhang, Sixin, Cun, Yann L, and Fergus, Rob · 2013
Earlier work this paper cites.
Fast dropout training
Wang, Sida I and Manning, Christopher D · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, Diederik and Ba, Jimmy · 2014
Cited alongside, same era.
Maeda, Shin-ichi · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, Danilo Jimenez, Mohamed, Shakir, and Wierstra, Daan · 2014
Cited alongside, same era.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, Nitish, Hinton, Geoffrey E, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan · 2014
Cited alongside, same era.
Tensorizing neural networks
Novikov, Alexander, Podoprikhin, Dmitrii, Osokin, Anton, and Vetrov, Dmitry P · 2015
Later among the works it cites.
Going deeper with convolutions
Szegedy, Christian, Liu, Wei, Jia, Yangqing, Sermanet, Pierre, Reed, Scott, Anguelov, Dragomir, Erhan, Dumitru, Vanhoucke, Vincent, and Rabinovich, Andrew · 2015
Later among the works it cites.
92.45 on cifar-10 in torch, 2015
Zagoruyko, Sergey · 2015
Later among the works it cites.
Ultimate tensorization: compressing convolutional and fc layers alike
Garipov, Timur, Podoprikhin, Dmitry, Novikov, Alexander, and Vetrov, Dmitry · 2016
Later among the works it cites.
Dynamic network surgery for efficient dnns
Guo, Yiwen, Yao, Anbang, and Chen, Yurong · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Doubly stochastic variational bayes for non-conjugate inference
Titsias, Michalis and Lázaro-Gredilla, Miguel · 2014
Cited alongside, same era.
Dropout as a bayesian approximation: Insights and applications
Gal, Yarin and Ghahramani, Zoubin · 2015
Cited alongside, same era.
Deep residual learning for image recognition
He, Kaiming, Zhang, Xiangyu, Ren, Shaoqing, and Sun, Jian · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, Sergey and Szegedy, Christian · 2015
Cited alongside, same era.
Variational dropout and the local reparameterization trick
Kingma, Diederik P, Salimans, Tim, and Welling, Max · 2015
Cited alongside, same era.
Fast convnets using group-wise brain damage
Lebedev, Vadim and Lempitsky, Victor · 2015
Cited alongside, same era.
Sparse convolutional neural networks
Liu, Baoyuan, Wang, Min, Foroosh, Hassan, Tappen, Marshall, and Pensky, Marianna · 2015
Cited alongside, same era.
Scardapane, Simone, Comminiello, Danilo, Hussain, Amir, and Uncini, Aurelio · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
Silver, David, Huang, Aja, Maddison, Chris J, Guez, Arthur, Sifre, Laurent, Van Den Driessche, George, Schrittwieser, Julian, Antonoglou, Ioannis, Panneershelvam, Veda, Lanctot, Marc, et al · 2016
Later among the works it cites.
How to train deep variational autoencoders and probabilistic ladder networks
Sønderby, Casper Kaae, Raiko, Tapani, Maaløe, Lars, Sønderby, Søren Kaae, and Winther, Ole · 2016
Later among the works it cites.
Srinivas, Suraj and Babu, R Venkatesh · 2016
Later among the works it cites.
Inception-v4, inception-resnet and the impact of residual connections on learning
Szegedy, Christian, Ioffe, Sergey, Vanhoucke, Vincent, and Alemi, Alex · 2016
Later among the works it cites.
Learning structured sparsity in deep neural networks
Wen, Wei, Wu, Chunpeng, Wang, Yandan, Chen, Yiran, and Li, Hai · 2016
Later among the works it cites.
Understanding deep learning requires rethinking generalization
Zhang, Chiyuan, Bengio, Samy, Hardt, Moritz, Recht, Benjamin, and Vinyals, Oriol · 2016
Later among the works it cites.
The power of sparsity in convolutional neural networks
Soravit Changpinyo, Mark Sandler, Andrey Zhmoginov · 2017
Closest in time.
Soft weight-sharing for neural network compression
Ullrich, Karen, Meeds, Edward, and Welling, Max · 2017
Closest in time.