Fetching the paper…
Reading the bibliography…
A method to increase the precision of feedforward networks is proposed.
H. Bourlard and Y. Kamp, “Auto-association by multilayer perceptrons and singular value decomposition,” Biological cybernetics , vol. 59, no. 4, pp. 291–294, 1988
1988
Earlier work this paper cites.
K. Hornik, M. Stinchcombe, and H. White, “Multilayer feedforward networks are universal approximators,” Neural networks , vol. 2, no. 5, pp. 359–366, 1989
1989
Earlier work this paper cites.
G. Cybenko, “Approximation by superpositions of a sigmoidal function,” Mathematics of Control, Signals, and Systems (MCSS) , vol. 2, no. 4, pp. 303–314, 1989
1989
Earlier work this paper cites.
K. Hornik, “Approximation capabilities of multilayer feedforward networks,” Neural networks , vol. 4, no. 2, pp. 251–257, 1991
1991
Earlier work this paper cites.
V. Kurkova, “Kolmogorov’s theorem and multilayer neural networks,” Neural networks , vol. 5, no. 3, pp. 501–506, 1992
1992
Earlier work this paper cites.
H. Drucker and Y. Le Cun, “Improving generalization performance using double backpropagation,” IEEE Transactions on Neural Networks , vol. 3, no. 6, pp. 991–997, 1992
1992
Earlier work this paper cites.
P. Cardaliaguet and G. Euvrard, “Approximation of a function and its derivative with a neural network,” Neural Networks , vol. 5, no. 2, pp. 207–220, 1992
1992
Earlier work this paper cites.
E. D. Sontag, “Feedback stabilization using two-hidden-layer nets,” IEEE Transactions on neural networks , vol. 3, no. 6, pp. 981–990, 1992
1992
Earlier work this paper cites.
A. R. Barron, “Universal approximation bounds for superpositions of a sigmoidal function,” IEEE Transactions on Information theory , vol. 39, no. 3, pp. 930–945, 1993
1993
Earlier work this paper cites.
M. Riedmiller and H. Braun, “A direct adaptive method for faster backpropagation learning: The rprop algorithm,” in Neural Networks, 1993., IEEE International Conference on . IEEE, 1993, pp. 586–591
1993
Earlier work this paper cites.
——, “Approximation and estimation bounds for artificial neural networks,” Machine Learning , vol. 14, no. 1, pp. 115–133, 1994
1994
Earlier work this paper cites.
A. Meade and A. A. Fernandez, “The numerical solution of linear ordinary differential equations by feedforward neural networks,” Mathematical and Computer Modelling , vol. 19, no. 12, pp. 1–25, 1994
1994
Earlier work this paper cites.
C. M. Bishop, Neural networks for pattern recognition . Oxford university press, 1995
1995
Earlier work this paper cites.
I. Lagaris, A. Likas, and D. Fotiadis, “Artificial neural network methods in quantum mechanics,” Computer Physics Communications , vol. 104, no. 1-3, pp. 1–14, 1997
1997
Earlier work this paper cites.
I. E. Lagaris, A. Likas, and D. I. Fotiadis, “Artificial neural networks for solving ordinary and partial differential equations,” IEEE Transactions on Neural Networks , vol. 9, no. 5, pp. 987–1000, 1998
1998
Earlier work this paper cites.
P. Simard, Y. LeCun, J. Denker, and B. Victorri, “Transformation invariance in pattern recognition—tangent distance and tangent propagation,” Neural networks: tricks of the trade , pp. 549–550, 1998
1998
Cited alongside, same era.
E. Basson and A. P. Engelbrecht, “Approximation of a function and its derivatives in feedforward neural networks,” in IJCNN’99. International Joint Conference on Neural Networks. Proceedings (Cat. No. 99CH36339) , vol. 1. IEEE, 1999, pp. 419–421
1999
Cited alongside, same era.
G. P. Zhang, “Neural networks for classification: a survey,” IEEE Transactions on Systems, Man, and Cybernetics, Part C (Applications and Reviews) , vol. 30, no. 4, pp. 451–462, 2000
2000
Cited alongside, same era.
S. He, K. Reif, and R. Unbehauen, “Multilayer neural networks for solving a class of partial differential equations,” Neural networks , vol. 13, no. 3, pp. 385–396, 2000
2000
Cited alongside, same era.
J. Bergstra, O. Breuleux, P. Lamblin, R. Pascanu, O. Delalleau, G. Desjardins, I. Goodfellow, A. Bergeron, Y. Bengio, and P. Kaelbling, “Theano: Deep learning on gpus with python,” 2011
2011
Later among the works it cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Later among the works it cites.
L. Deng and X. Li, “Machine learning paradigms for speech recognition: An overview,” IEEE Transactions on Audio, Speech, and Language Processing , vol. 21, no. 5, pp. 1060–1089, 2013
2013
Later among the works it cites.
L. Wan, M. Zeiler, S. Zhang, Y. Le Cun, and R. Fergus, “Regularization of neural networks using dropconnect,” in International Conference on Machine Learning , 2013, pp. 1058–1066
2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
G. W. Flake and B. A. Pearlmutter, “Differentiating functions of the jacobian with respect to the weights,” in Advances in Neural Information Processing Systems , 2000, pp. 435–441
2000
Cited alongside, same era.
R. Murray-Smith and B. A. Pearlmutter, “Transformations of gaussian process priors,” in Deterministic and Statistical Methods in Machine Learning . Springer, 2005, pp. 110–123
2005
Cited alongside, same era.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in Proceedings of the 23rd international conference on Machine learning . ACM, 2006, pp. 369–376
2006
Cited alongside, same era.
A. Malek and R. S. Beidokhti, “Numerical solution for high order differential equations using a hybrid neural network optimization method,” Applied Mathematics and Computation , vol. 183, no. 1, pp. 260–271, 2006
2006
Cited alongside, same era.
M. Hardy, “Combinatorics of partial derivatives,” the electronic journal of combinatorics , vol. 13, no. 1, p. 1, 2006
2006
Cited alongside, same era.
A. Griewank and A. Walther, Evaluating derivatives: principles and techniques of algorithmic differentiation . SIAM, 2008
2008
Cited alongside, same era.
C. Nvidia, “Cublas library,” NVIDIA Corporation, Santa Clara, California , vol. 15, no. 27, p. 31, 2008
2008
Cited alongside, same era.
Y. Bengio et al. , “Learning deep architectures for ai,” Foundations and trends® in Machine Learning , vol. 2, no. 1, pp. 1–127, 2009
2009
Cited alongside, same era.
2013
Later among the works it cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: A simple way to prevent neural networks from overfitting,” The Journal of Machine Learning Research , vol. 15, no. 1, pp. 1929–1958, 2014
2014
Later among the works it cites.
2015
Later among the works it cites.
D. Mishkin and J. Matas, “All you need is a good init,” arXiv preprint arXiv:1511.06422 , 2015
2015
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 1026–1034
2015
Later among the works it cites.
2016
Later among the works it cites.
S. Petridis and M. Pantic, “Deep complementary bottleneck features for visual speech recognition,” in 2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2016, pp. 2304–2308
2016
Later among the works it cites.
2016
Later among the works it cites.
2017
Closest in time.
J. Berg and K. Nyström, “A unified deep artificial neural network approach to partial differential equations in complex geometries,” Neurocomputing , vol. 317, pp. 28–41, 2018
2018
Closest in time.
X. Saint Raymond, Elementary introduction to the theory of pseudodifferential operators . Routledge, 2018
2018
Closest in time.