“Positive definite functions on spheres,”
I. J. Schoenberg, · 1942
Earlier work this paper cites.
“Orthogonal random features,”
Felix Xinnan Yu, Ananda Theertha Suresh, Krzysztof Choromanski, Daniel Holtmannrice, and Sanjiv Kumar, · 1983
Earlier work this paper cites.
Correlation Theory of Stationary and Related Random Functions
Akiva M Yaglom, · 1987
Earlier work this paper cites.
Spin glass theory and beyond: An Introduction to the Replica Method and Its Applications
Marc Mézard, Giorgio Parisi, and Miguel Angel Virasoro, · 1987
Earlier work this paper cites.
Random number generation and quasi-Monte Carlo methods
Harald Niederreiter, · 1992
Earlier work this paper cites.
Practical numerical integration
Gwynne Evans, · 1993
Earlier work this paper cites.
“Computing with infinite networks,”
Christopher K.I. Williams, · 1997
Earlier work this paper cites.
“Stochastic integration rules for infinite regions,”
Alan Genz and John Monahan, · 1998
Earlier work this paper cites.
“A stochastic algorithm for high-dimensional integrals over unbounded regions with gaussian weight,”
Alan Genz and John Monahan, · 1999
Earlier work this paper cites.
“Sparse greedy matrix approximation for machine learning,”
Alex J. Smola and Bernhard Schölkopf, · 2000
Earlier work this paper cites.
“Using the Nyström method to speed up kernel machines,”
Christopher K.I. Williams and Matthias Seeger, · 2001
Earlier work this paper cites.
“Regularization with dot-product kernels,”
Alex J. Smola, Zoltan L. Ovari, and Robert C. Williamson, · 2001
Earlier work this paper cites.
“Classes of kernels for machine learning: a statistics perspective,”
Marc G. Genton, · 2001
Earlier work this paper cites.
Least Squares Support Vector Machines
Johan A.K. Suykens, Tony Van Gestel, Jos De Brabanter, Bart De Moor, and Joos Vandewalle, · 2002
Earlier work this paper cites.
“On the eigenspectrum of the gram matrix and its relationship to the operator eigenspectrum,”
John Shawe-Taylor, Chris Williams, Nello Cristianini, and Jaz Kandola, · 2002
Earlier work this paper cites.
“On the mathematical foundations of learning,”
Felipe Cucker and Steve Smale, · 2002
Earlier work this paper cites.
Learning with kernels: support vector machines, regularization, optimization, and beyond
Bernhard Schölkopf and Alexander J. Smola, · 2003
Earlier work this paper cites.
Harmonic Analysis and the Theory of Probability
Salomon Bochner, · 2005
Earlier work this paper cites.
“Learning bounds for kernel regression using effective data dimensionality,”
Tong Zhang, · 2005
Earlier work this paper cites.
“Local rademacher complexities,”
Peter L Bartlett, Olivier Bousquet, and Shahar Mendelson, · 2005
Earlier work this paper cites.
Spherical harmonics
Claus Müller, · 2006
Earlier work this paper cites.
“Extreme learning machine: theory and applications,”
Guang-Bin Huang, Qin-Yu Zhu, and Chee-Kheong Siew, · 2006
Earlier work this paper cites.
“Random features for large-scale kernel machines,”
Ali Rahimi and Benjamin Recht, · 2007
Earlier work this paper cites.
Learning theory: an approximation theory viewpoint
Felipe Cucker and Dingxuan Zhou, · 2007
Earlier work this paper cites.
“Optimal rates for the regularized least-squares algorithm,”
Andrea Caponnetto and Ernesto De Vito, · 2007
Earlier work this paper cites.
“Learning theory estimates via integral operators and their approximations,”
Steve Smale and Ding-Xuan Zhou, · 2007
Earlier work this paper cites.
“Fast rates for support vector machines using Gaussian kernels,”
Ingo Steinwart and Clint Scovel, · 2007
Earlier work this paper cites.
“Training invariant support vector machines using selective sampling,”
Gaëlle Loosli, Stéphane Canu, and Léon Bottou, · 2007
Earlier work this paper cites.
“Likelihood approximation by numerical integration on sparse grids,”
Florian Heiss and Viktor Winschel, · 2008
Earlier work this paper cites.
“Uniform approximation of functions with random bases,”
Ali Rahimi and Benjamin Recht, · 2008
Earlier work this paper cites.
Support Vector Machines
Ingo Steinwart and Christmann Andreas, · 2008
Earlier work this paper cites.
“Weighted sums of random kitchen sinks: Replacing minimization with randomization in learning,”
Ali Rahimi and Benjamin Recht, · 2009
Earlier work this paper cites.
“Kernel methods for deep learning,”
Youngmin Cho and Lawrence K Saul, · 2009
Earlier work this paper cites.
“Learning multiple layers of features from tiny images,”
Alex Krizhevsky and Geoffrey Hinton, · 2009
Earlier work this paper cites.
“Imagenet: A large-scale hierarchical image database,”
Jia Deng, Wei Dong, Richard Socher, Li Jia Li, Kai Li, and Fei Fei Li, · 2009
Earlier work this paper cites.
“Random Fourier approximations for skewed multiplicative histogram kernels,”
Fuxin Li, Catalin Ionescu, and Cristian Sminchisescu, · 2010
Earlier work this paper cites.
“Two-stage learning kernel algorithms,”
Corinna Cortes, Mehryar Mohri, and Afshin Rostamizadeh, · 2010
Earlier work this paper cites.
“Optimal learning rates for kernel conjugate gradient regression,”
Gilles Blanchard and Nicole Krämer, · 2010
Earlier work this paper cites.
Oracle Inequalities in Empirical Risk Minimization and Sparse Recovery Problems
Vladimir Koltchinskii, · 2011
Earlier work this paper cites.
“Random feature maps for dot product kernels,”
Purushottam Kar and Harish Karnick, · 2012
Earlier work this paper cites.
“Efficient additive kernels via explicit feature maps,”
Andrea Vedaldi and Andrew Zisserman, · 2012
Earlier work this paper cites.
“Nyström method vs random Fourier features: a theoretical and empirical comparison,”
Tianbao Yang, Yu Feng Li, Mehrdad Mahdavi, Rong Jin, and Zhi Hua Zhou, · 2012
Earlier work this paper cites.
Topics in random matrix theory
Terence Tao, · 2012
Earlier work this paper cites.
“Divide and conquer kernel ridge regression,”
Yuchen Zhang, John Duchi, and Martin Wainwright, · 2013
Earlier work this paper cites.
“FastFood—approximating kernel expansions in loglinear time,”
Quoc Le, Tamás Sarlós, and Alex J. Smola, · 2013
Earlier work this paper cites.
“Fast and scalable polynomial kernels via explicit feature maps,”
Ninh Pham and Rasmus Pagh, · 2013
Earlier work this paper cites.
“Gaussian process kernels for pattern discovery and extrapolation,”
Andrew Gordon Wilson and Ryan Prescott Adams, · 2013
Earlier work this paper cites.
“Sharp analysis of low-rank kernel matrix approximations,”
Francis Bach, · 2013
Earlier work this paper cites.
“A divide-and-conquer solver for kernel support vector machines,”
Cho-Jui Hsieh, Si Si, and Inderjit Dhillon, · 2014
Earlier work this paper cites.
“Randomized nonlinear component analysis,”
David Lopez-Paz, Suvrit Sra, Alex J. Smola, Zoubin Ghahramani, and Bernhard Schölkopf, · 2014
Earlier work this paper cites.
“Quasi-Monte Carlo feature maps for shift-invariant kernels,”
Jiyan Yang, Vikas Sindhwani, Haim Avron, and Michael Mahoney, · 2014
Earlier work this paper cites.
“Scalable kernel methods via doubly stochastic gradients,”
Bo Dai, Bo Xie, Niao He, Yingyu Liang, Anant Raj, Maria-Florina F Balcan, and Le Song, · 2014
Earlier work this paper cites.
“Subspace embeddings for the polynomial kernel,”
Haim Avron, Huy Nguyen, and David Woodruff, · 2014
Earlier work this paper cites.
“Compact random feature maps,”
Raffay Hamid, Ying Xiao, Alex Gittens, and Dennis Decoste, · 2014
Earlier work this paper cites.
“On the error of random Fourier features,”
Danica J. Sutherland and Jeff Schneider, · 2015
Earlier work this paper cites.
“Spherical random features for polynomial kernels,”
Jeffrey Pennington, Felix Xinnan X. Yu, and Sanjiv Kumar, · 2015
Earlier work this paper cites.
“Random feature mapping with signed circulant matrix projection,”
Chang Feng, Qinghua Hu, and Shizhong Liao, · 2015
Earlier work this paper cites.
“Optimal rates for random Fourier features,”
Bharath K. Sriperumbudur and Zoltán Szabó, · 2015
Earlier work this paper cites.
“Compact nonlinear maps and circulant extensions,”
Original
Felix X. Yu, Sanjiv Kumar, Henry Rowley, and Shih Fu Chang, · 2015
Earlier work this paper cites.
“À la carte–learning fast kernels,”
Zichao Yang, Andrew Wilson, Alex J. Smola, and Le Song, · 2015
Earlier work this paper cites.