Fetching the paper…
Reading the bibliography…
We propose a new regularization method based on virtual adversarial loss: a new measure of local smoothness of the conditional label distribution given input.
Solutions of ill-posed problems
Andrej N Tikhonov and Vasiliy Y Arsenin · 1977
Earlier work this paper cites.
Spline models for observational data
Grace Wahba · 1990
Earlier work this paper cites.
Regularization using jittered training data
Russell Reed, Seho Oh, and RJ Marks · 1992
Earlier work this paper cites.
Training with noise is equivalent to Tikhonov regularization
Christopher M Bishop · 1995
Earlier work this paper cites.
Information theory and an extension of the maximum likelihood principle
Hirotugu Akaike · 1998
Earlier work this paper cites.
Eigenvalue computation in the 20th century
Gene H Golub and Henk A van der Vorst · 2000
Earlier work this paper cites.
Learning from labeled and unlabeled data with label propagation
Xiaojin Zhu and Zoubin Ghahramani · 2002
Earlier work this paper cites.
Semi-supervised learning by entropy minimization
Yves Grandvalet and Yoshua Bengio · 2004
Earlier work this paper cites.
Pattern Recognition and Machine Learning
Christopher M Bishop · 2006
Earlier work this paper cites.
Large scale transductive SVMs
Ronan Collobert, Fabian Sinz, Jason Weston, and Léon Bottou · 2006
Earlier work this paper cites.
What is the best multi-stage architecture for object recognition?
Kevin Jarrett, Koray Kavukcuoglu, Marc’Aurelio Ranzato, and Yann LeCun · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton · 2009
Earlier work this paper cites.
Algebraic geometry and statistical learning theory
Sumio Watanabe · 2009
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
Vinod Nair and Geoffrey E Hinton · 2010
Earlier work this paper cites.
Deep sparse rectifier neural networks
Xavier Glorot, Antoine Bordes, and Yoshua Bengio · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Yuval Netzer, Tao Wang, Adam Coates, Alessandro Bissacco, Bo Wu, and Andrew Y Ng · 2011
Earlier work this paper cites.
Mathematical methods of classical mechanics
Vladimir Igorevich Arnol’d · 2013
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Andrew L Maas, Awni Y Hannun, and Andrew Y Ng · 2013
Cited alongside, same era.
Dropout training as adaptive regularization
Stefan Wager, Sida Wang, and Percy S Liang · 2013
Cited alongside, same era.
Learning with pseudo-ensembles
Philip Bachman, Ouais Alsharif, and Doina Precup · 2014
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Semi-supervised learning with deep generative models
Diederik Kingma, Shakir Mohamed, Danilo Jimenez Rezende, and Max Welling · 2014
Cited alongside, same era.
Network in network
Min Lin, Qiang Chen, and Shuicheng Yan · 2014
Cited alongside, same era.
Striving for simplicity: The all convolutional net
Jost Tobias Springenberg, Alexey Dosovitskiy, Thomas Brox, and Martin Riedmiller · 2015
Later among the works it cites.
Rupesh Kumar Srivastava, Klaus Greff, and Jürgen Schmidhuber · 2015
Later among the works it cites.
Chainer: a next-generation open source framework for deep learning
Seiya Tokui, Kenta Oono, Shohei Hido, and Justin Clayton · 2015
Later among the works it cites.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Martın Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, et al · 2016
Later among the works it cites.
Understanding regularization by virtual adversarial training, ladder networks and others
Mudassar Abbas, Jyri Kivinen, and Tapani Raiko · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shin-ichi Maeda · 2014
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Cited alongside, same era.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2014
Cited alongside, same era.
Explaining and harnessing adversarial examples
Ian Goodfellow, Jonathon Shlens, and Christian Szegedy · 2015
Cited alongside, same era.
Towards deep neural network architectures robust to adversarial examples
Shixiang Gu and Luca Rigazio · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Yarin Gal and Zoubin Ghahramani · 2016
Later among the works it cites.
Deep Learning
Ian Goodfellow, Yoshua Bengio, and Aaron Courville · 2016
Later among the works it cites.
Identity mappings in deep residual networks
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Later among the works it cites.
Auxiliary deep generative models
Lars Maaløe, Casper Kaae Sønderby, Søren Kaae Sønderby, and Ole Winther · 2016
Later among the works it cites.
Distributional smoothing with virtual adversarial training
Takeru Miyato, Shin - · 2016
Later among the works it cites.
Regularization with stochastic transformations and perturbations for deep semi-supervised learning
Mehdi Sajjadi, Mehran Javanmardi, and Tolga Tasdizen · 2016
Later among the works it cites.
Improved techniques for training GANs
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen · 2016
Later among the works it cites.
Theano: A Python framework for fast computation of mathematical expressions
Theano Development Team · 2016
Later among the works it cites.
Stacked what-where auto-encoders
Junbo Zhao, Michael Mathieu, Ross Goroshin, and Yann Lecun · 2016
Later among the works it cites.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, and Kilian Q Weinberger · 2017
Closest in time.
Temporal ensembling for semi-supervised learning
Samuli Laine and Timo Aila · 2017
Closest in time.