Fetching the paper…
Reading the bibliography…
In this paper we provide a finite-sample and an infinite-sample representer theorem for the concatenation of (linear combinations of) kernel functions of reproducing kernel Hilbert spaces.
N. Aronszajn, Theory of reproducing kernels , Transactions of the American Mathematical Society 68
1950
Earlier work this paper cites.
G. Kimeldorf and G. Wahba, A correspondence between Bayesian estimation on stochastic processes and smoothing by splines , The Annals of Mathematical Statistics 41
1970
Earlier work this paper cites.
G. Cybenko, Approximations by superpositions of sigmoidal functions , Mathematics of Control, Signals, and Systems 2
1989
Earlier work this paper cites.
K. Hornik, Approximation capabilities of multilayer feedforward networks , Neural Networks 4
1991
Earlier work this paper cites.
B. Schölkopf and A. Smola, Learning with Kernels – Support Vector Machines, Regularization, Optimization, and Beyond , The MIT Press – Cambridge, Massachusetts, 2002
2002
Earlier work this paper cites.
F. Bach, G. Lanckriet, and M. Jordan, Multiple kernel learning, conic duality, and the SMO algorithm , Proceedings of the 21st International Conference on Machine Learning, 2004, pp. 1–9
2004
Earlier work this paper cites.
H.-J. Bungartz and M. Griebel, Sparse grids , Acta Numerica 13
2004
Earlier work this paper cites.
C. Micchelli and M. Pontil, On learning vector-valued functions , Neural Computation 17
2005
Earlier work this paper cites.
H. Wendland, Scattered Data Approximation , Cambridge Monographs on Applied and Computational Mathematics, Cambridge University Press, 2005
2005
Earlier work this paper cites.
I. Steinwart and A. Christmann, Support vector machines , Springer, New York, 2008
2008
Earlier work this paper cites.
Y. Cho and L. Saul, Kernel methods for deep learning , Advances in Neural Information Processing Systems (Y. Bengio, D. Schuurmans, J. Lafferty, C. Williams, and A. Culotta, eds.), vol. 22, Curran Associates, Inc., 2009, pp. 342–350
2009
Cited alongside, same era.
S. de Marchi and R. Schaback, Stability of kernel-based interpolation , Adv. Comput. Math. 32
2010
Cited alongside, same era.
F. Dinuzzo, Learning functions with kernel methods , Ph.D. thesis, University of Pavia, Pavia, Italy, 2011
2011
Cited alongside, same era.
G. Fasshauer and Q. Ye, Reproducing kernels of generalized Sobolev spaces via a Green function approach with distributional operators , Numerische Mathematik 119
2011
Cited alongside, same era.
M. Gönen and E. Alpaydin, Multiple kernel learning algorithms , JMLR 12
2011
Cited alongside, same era.
I. Goodfellow, Y. Bengio, and A. Courville, Deep Learning , MIT Press, 2016
2016
Later among the works it cites.
A. Hinrichs, L. Markhasin, J. Oettershagen, and T. Ullrich, Optimal quasi-Monte Carlo rules on higher order digital nets for the numerical integration of multivariate periodic functions , Numerische Mathematik 134
2016
Later among the works it cites.
I. Rebai, Y. Benayed, and W. Mahdi, Deep multilayer multiple kernel learning , Neural Computing and Applications 27
2016
Later among the works it cites.
S. Reddi, S. Sra, B. Poczos, and A. Smola, Proximal stochastic methods for nonsmooth nonconvex finite-sum optimization , Advances in Neural Information Processing Systems 29 (D. D. Lee, M. Sugiyama, U. V. Luxburg, I. Guyon, and R. Garnett, eds.), Curran Associates, Inc., 2016, pp. 1145–1153
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Zhuang, I. Tsang, and S. Hoi, Two-layer multiple kernel learning , Proceedings of the 14th International Conference on Artificial Intelligence and Statistics, 2011, pp. 909–917
2011
Cited alongside, same era.
A. Damianou and N. Lawrence, Deep Gaussian processes , Proceedings of the 16th International Conference on Artificial Intelligence and Statistics, 2013, pp. 207–215
2013
Cited alongside, same era.
E. Strobl and S. Visweswaran, Deep multiple kernel learning , Proceedings of the 12th International Conference on Machine Learning and Applications, 2013, pp. 414–417
2013
Cited alongside, same era.
M. Griebel and H. Harbrecht, Approximation of bi-variate functions: singular value decomposition versus sparse grids , IMA J. Numer. Anal. 34
2014
Cited alongside, same era.
A. Wilson, Z. Hu, R. Salakhutdinov, and E. Xing, Deep kernel learning , Proceedings of the 19th International Conference on Artificial Intelligence and Statistics, 2016, pp. 370–378
2016
Later among the works it cites.
B. Bohn and M. Griebel, Error estimates for multivariate regression on discretized function spaces , SIAM Journal on Numerical Analysis 55
2017
Closest in time.
H. Mhaskar, Q. Liao, and T. Poggio, When and why are deep networks better than shallow ones? , Proceedings of the 31st AAAI Conference on Artificial Intelligence, 2017, pp. 2343–2349
2017
Closest in time.
G. Montavon, S. Lapuschkin, A. Binder, W. Samek, and K.-R. Müller, Explaining nonlinear classification decisions with deep Taylor decomposition , Pattern Recognition 65
2017
Closest in time.
S. Mallat, Understanding deep convolutional networks , Phil. Trans. R. Soc. A 374
2065
Closest in time.