Fetching the paper…
Reading the bibliography…
Activation functions play an important role in training artificial neural networks.
G. Cybenko, “Approximation by superpositions of a sigmoidal function,” Mathematics of Control, Signals, and Systems , vol. 2, no. 4, pp. 303–314, dec 1989
1989
Earlier work this paper cites.
D. J. MacKay, “A practical bayesian framework for backpropagation networks,” Neural computation , vol. 4, no. 3, pp. 448–472, 1992
1992
Earlier work this paper cites.
A. Murray and P. Edwards, “Synaptic weight noise during multilayer perceptron training: fault tolerance and training improvements,” IEEE Transactions on Neural Networks , vol. 4, no. 4, pp. 722–725, jul 1993
1993
Earlier work this paper cites.
G. Hinton and D. Van Camp, “Keeping neural networks simple by minimizing the description length of the weights,” in in Proc. of the 6th Ann. ACM Conf. on Computational Learning Theory . Citeseer, 1993
1993
Earlier work this paper cites.
C. M. Bishop, “Training with noise is equivalent to tikhonov regularization,” Neural Computation , vol. 7, no. 1, pp. 108–116, jan 1995
1995
Earlier work this paper cites.
G. An, “The effects of adding noise during backpropagation training on a generalization performance,” Neural Computation , vol. 8, no. 3, pp. 643–674, apr 1996
1996
Earlier work this paper cites.
M. S. Lewicki, “A review of methods for spike sorting: the detection and classification of neural action potentials,” Network: Computation in Neural Systems , vol. 9, no. 4, pp. R53–R78, 1998
1998
Earlier work this paper cites.
H. Inayoshi and T. Kurita, “Improved generalization by adding both auto-association and hidden-layer-noise to neural-network-based-classifiers,” in IEEE Workshop on Machine Learning for Signal Processing , 2005
2005
Earlier work this paper cites.
P. Vincent, H. Larochelle, Y. Bengio, and P.-A. Manzagol, “Extracting and composing robust features with denoising autoencoders,” in ACM International Conference on Machine Learning , 2008, pp. 1096–1103
2008
Earlier work this paper cites.
V. Nair and G. E. Hinton, “Rectified linear units improve restricted boltzmann machines,” in International Conference on Machine Learning , 2010, pp. 807–814
2010
Earlier work this paper cites.
A. Vehbi Olgac and B. Karlik, “Performance analysis of various activation functions in generalized MLP architectures of neural networks,” International Journal of Artificial Intelligence And Expert Systems , vol. 1, pp. 111–122, 02 2011
2011
Earlier work this paper cites.
A. Graves, “Practical variational inference for neural networks,” in Advances in Neural Information Processing Systems , J. Shawe-Taylor, R. S. Zemel, P. L. Bartlett, F. Pereira, and K. Q. Weinberger, Eds., 2011, pp. 2348–2356
2011
Earlier work this paper cites.
A. Graves, “Practical variational inference for neural networks,” in Advances in neural information processing systems , 2011, pp. 2348–2356
2011
Earlier work this paper cites.
A. Coates, A. Ng, and H. Lee, “An analysis of single-layer networks in unsupervised feature learning,” in International Conference on Artificial Intelligence and Statistics , 2011, pp. 215–223
2011
Earlier work this paper cites.
A. L. Maas, R. E. Daly, P. T. Pham, D. Huang, A. Y. Ng, and C. Potts, “Learning word vectors for sentiment analysis,” in Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics: Human Language Technologies . Portland, Oregon, USA: Association for Computational Linguistics, June 2011, pp. 142–150. [Online]. Available: http://www.aclweb.org/anthology/P11-1015
2011
Earlier work this paper cites.
2012
Cited alongside, same era.
K. Audhkhasi, O. Osoba, and B. Kosko, “Noise benefits in backpropagation and deep bidirectional pre-training,” in International Joint Conference on Neural Networks , aug 2013
2013
Cited alongside, same era.
2013
Cited alongside, same era.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,” Journal of Machine Learning Research , vol. 15, no. 1, pp. 1929–1958, 2014
2014
Cited alongside, same era.
C. Gulcehre, M. Moczulski, M. Denil, and Y. Bengio, “Noisy activation functions,” in International Conference on Machine Learning , vol. 48, jun 2016, pp. 3059–3068
2016
Later among the works it cites.
2016
Later among the works it cites.
2017
Later among the works it cites.
L. Trottier, P. Gigu, B. Chaib-draa et al. , “Parametric exponential linear unit for deep convolutional neural networks,” in IEEE International Conference on Machine Learning and Applications , 2017, pp. 207–214
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
J. Schmidhuber, “Deep learning in neural networks: An overview,” Neural Networks , vol. 61, pp. 85–117, 2015
2015
Cited alongside, same era.
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on ImageNet classification,” in International Conference on Computer Vision , dec 2015
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2015
Cited alongside, same era.
C. Blundell, J. Cornebise, K. Kavukcuoglu, and D. Wierstra, “Weight uncertainty in neural network,” in International Conference on Machine Learning , vol. 37, 2015, pp. 1613–1622
2015
Cited alongside, same era.
J. M. Hernández-Lobato and R. Adams, “Probabilistic backpropagation for scalable learning of bayesian neural networks,” in International Conference on Machine Learning , 2015, pp. 1861–1869
2015
Cited alongside, same era.
A. Kendall and Y. Gal, “What uncertainties do we need in bayesian deep learning for computer vision?” in Advances in neural information processing systems , 2017, pp. 5574–5584
2017
Later among the works it cites.
B. Lakshminarayanan, A. Pritzel, and C. Blundell, “Simple and scalable predictive uncertainty estimation using deep ensembles,” in Advances in neural information processing systems , 2017, pp. 6402–6413
2017
Later among the works it cites.
2017
Later among the works it cites.
2018
Later among the works it cites.
X. Liu, M. Cheng, H. Zhang, and C.-J. Hsieh, “Towards robust neural networks via random self-ensemble,” in European Conference on Computer Vision , 2018, pp. 369–385
2018
Later among the works it cites.
Y. Kwon, J.-H. Won, B. Kim, and M. Paik, “Uncertainty quantification using bayesian neural networks in classification: Application to ischemic stroke lesion segmentation,” Computational Statistics and Data Analysis , 04 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.