Fetching the paper…
Reading the bibliography…
We propose to optimize the activation functions of a deep neural network by adding a corresponding functional regularization to the cost function.
Theory of reproducing kernels
Nachman Aronszajn · 1950
Earlier work this paper cites.
Spline functions and the problem of graduation
I. J. Schoenberg · 1964
Earlier work this paper cites.
On splines and their minimum properties
C. de Boor and R. E. Lynch · 1966
Earlier work this paper cites.
Théorie des Distributions
Laurent Schwartz · 1966
Earlier work this paper cites.
Some results on Tchebycheffian spline functions
George Kimeldorf and Grace Wahba · 1971
Earlier work this paper cites.
Spline solutions to L 1 L_{1} extremal problems in one and several variables
SD Fisher and JW Jerome · 1975
Earlier work this paper cites.
Splines and Variational Methods
P.M. Prenter · 1975
Earlier work this paper cites.
A Practical Guide to Splines
C. de Boor · 1978
Earlier work this paper cites.
Methods of Modern Mathematical Physics. Vol. 1: Functional Analysis , volume 1
Michael Reed and Barry Simon · 1980
Earlier work this paper cites.
Spline Functions: Basic Theory
L.L. Schumaker · 1981
Earlier work this paper cites.
Interpolation of scattered data: Distance matrices and conditionally positive definite functions
Charles A Micchelli · 1986
Earlier work this paper cites.
Learning representations by back-propagating errors
David E Rumelhart, Geoffrey E Hinton, and Ronald J Williams · 1986
Earlier work this paper cites.
Real and Complex Analysis
Walter Rudin · 1987
Earlier work this paper cites.
Regularization algorithms for learning that are equivalent to multilayer networks
Tomaso Poggio and Federico Girosi · 1990
Earlier work this paper cites.
Spline Models for Observational Data
G. Wahba · 1990
Earlier work this paper cites.
Multi-layer perceptrons with B-spline receptive field functions
Stephen H Lane, Marshall Flax, David Handelman, and Jack Gelfand · 1991
Earlier work this paper cites.
Functional Analysis
Walter Rudin · 1991
Earlier work this paper cites.
Locally adaptive regression splines
E. Mammen and S. van de Geer · 1997
Earlier work this paper cites.
Comparing support vector machines with Gaussian kernels to radial basis function classifiers
B. Schölkopf, Kah-Kay Sung, C. J. C. Burges, F. Girosi, P. Niyogi, T. Poggio, and V. Vapnik · 1997
Earlier work this paper cites.
Learning and approximation capabilities of adaptive spline activation function neural networks
Lorenzo Vecci, Francesco Piazza, and Aurelio Uncini · 1998
Cited alongside, same era.
Multilayer feedforward networks with adaptive spline activation function
Stefano Guarnieri, Francesco Piazza, and Aurelio Uncini · 1999
Cited alongside, same era.
Region configurations for realizability of lattice piecewise-linear models
J.M. Tarela and M.V. Martinez · 1999
Cited alongside, same era.
Splines: A perfect fit for signal and image processing
M. Unser · 1999
Cited alongside, same era.
Regularization networks and support vector machines
Theodoros Evgeniou, Massimiliano Pontil, and Tomaso Poggio · 2000
Cited alongside, same era.
A generalized representer theorem
Bernhard Schölkopf, Ralf Herbrich, and Alex J. Smola · 2001
Cited alongside, same era.
A Mathematical Introduction to Compressive Sensing
Simon Foucart and Holger Rauhut · 2013
Later among the works it cites.
Maxout networks
Ian J Goodfellow, David Warde-Farley, Mehdi Mirza, Aaron Courville, and Yoshua Bengio · 2013
Later among the works it cites.
The Nature of Statistical Learning Theory
Vladimir Vapnik · 2013
Later among the works it cites.
On the number of linear regions of deep neural networks
Guido F Montufar, Razvan Pascanu, Kyunghyun Cho, and Yoshua Bengio · 2014
Later among the works it cites.
Learning activation functions to improve deep neural networks
Forest Agostinelli, Matthew Hoffman, Peter Sadowski, and Pierre Baldi · 2015
Later among the works it cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning with Kernels: Support Vector Machines, Regularization, Optimization, and Beyond
Bernhard Schölkopf and Alexander J Smola · 2002
Cited alongside, same era.
The mathematics of learning: Dealing with data
Tomaso Poggio and Steve Smale · 2003
Cited alongside, same era.
Generalization of hinging hyperplanes
Shuning Wang and Xusheng Sun · 2005
Cited alongside, same era.
Scattered Data Approximations
H. Wendland · 2005
Cited alongside, same era.
Pattern Recognition and Machine Learning
Christopher M. Bishop · 2006
Cited alongside, same era.
For most large underdetermined systems of linear equations the minimal ℓ 1 \ell_{1} -norm solution is also the sparsest solution
D. L. Donoho · 2006
Cited alongside, same era.
Later among the works it cites.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Later among the works it cites.
Notes on hierarchical splines, DCLNs and i-theory
Tomaso Poggio, Lorenzo Rosasco, Amnon Shashua, Nadav Cohen, and Fabio Anselmi · 2015
Later among the works it cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Later among the works it cites.
Deep learning in neural networks: An overview
Jürgen Schmidhuber · 2015
Later among the works it cites.
Understanding deep neural networks with rectified linear units
Raman Arora, Amitabh Basu, Poorya Mianjy, and Anirbit Mukherjee · 2016
Later among the works it cites.
Deep Learning , volume 1
Ian Goodfellow, Yoshua Bengio, and Aaron Courville · 2016
Later among the works it cites.
Representer theorems for sparsity-promoting ℓ 1 \ell_{1} regularization
M. Unser, J. Fageot, and H. Gupta · 2016
Later among the works it cites.
Convnets with smooth adaptive activation functions for regression
Le Hou, Dimitris Samaras, Tahsin Kurc, Yi Gao, and Joel Saltz · 2017
Later among the works it cites.
Splines are universal solutions of linear inverse problems with generalized-TV regularization
M. Unser, J. Fageot, and J. P. Ward · 2017
Later among the works it cites.
A representer theorem for deep kernel learning
Bastian Bohn, Michael Griebel, and Christian Rieger · 2018
Closest in time.
Continuous-domain solutions of linear inverse problems with Tikhonov versus
H. Gupta, J. Fageot, and M. Unser · 2018
Closest in time.
The functions of deep learning
Gil Strang · 2018
Closest in time.