Fetching the paper…
Reading the bibliography…
Neural kernels have drastically increased performance on diverse and nonstandard data modalities but require significantly more compute, which previously limited their application to smaller datasets.
Priors for infinite networks (tech. rep. no. crg-tr-94-1)
Radford M. Neal · 1994
Earlier work this paper cites.
An introduction to the conjugate gradient method without the agonizing pain, 1994
Jonathan Richard Shewchuk et al · 1994
Earlier work this paper cites.
Incorporating invariances in support vector learning machines
Bernhard Schölkopf, Chris Burges, and Vladimir Vapnik · 1996
Earlier work this paper cites.
Computing with infinite networks
Christopher Williams · 1996
Earlier work this paper cites.
Incorporating prior information in machine learning by creating virtual examples
Partha Niyogi, Federico Girosi, and Tomaso Poggio · 1998
Earlier work this paper cites.
A Bayesian committee machine
V. Tresp · 2000
Earlier work this paper cites.
Gaussian processes for machine learning , volume 2
Christopher KI Williams and Carl Edward Rasmussen · 2006
Earlier work this paper cites.
Random features for large-scale kernel machines
Ali Rahimi and Benjamin Recht · 2007
Earlier work this paper cites.
Gaussian processes for big data
James Hensman, Nicolò Fusi, and Neil D. Lawrence · 2013
Earlier work this paper cites.
Tiny imagenet visual recognition challenge
Ya Le and Xuan Yang · 2015
Earlier work this paper cites.
Kernel interpolation for scalable structured gaussian processes (kiss-gp)
Andrew Gordon Wilson and Hannes Nickisch · 2015
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Thomas N Kipf and Max Welling · 2016
Earlier work this paper cites.
Google vizier: A service for black-box optimization
Daniel Golovin, Benjamin Solnik, Subhodeep Moitra, Greg Kochanski, John Karro, and David Sculley · 2017
Earlier work this paper cites.
Deep learning scaling is predictable, empirically
Joel Hestness, Sharan Narang, Newsha Ardalani, Gregory Diamos, Heewoo Jun, Hassan Kianinejad, Md Patwary, Mostofa Ali, Yang Yang, and Yanqi Zhou · 2017
Earlier work this paper cites.
Diving into the shallows: a computational perspective on large-scale shallow learning
Siyuan Ma and Mikhail Belkin · 2017
Earlier work this paper cites.
Falkon: An optimal large scale kernel method
Alessandro Rudi, Luigi Carratino, and Lorenzo Rosasco · 2017
Earlier work this paper cites.
JAX: composable transformations of Python+NumPy programs, 2018
James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, George Necula, Adam Paszke, Jake VanderPlas, Skye Wanderman-Milne, and Qiao Zhang · 2018
Earlier work this paper cites.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clement Hongler · 2018
Earlier work this paper cites.
Deep neural networks as gaussian processes
Jaehoon Lee, Yasaman Bahri, Roman Novak, Sam Schoenholz, Jeffrey Pennington, and Jascha Sohl-dickstein · 2018
Earlier work this paper cites.
Gaussian process behaviour in wide deep neural networks
Alexander G. de G. Matthews, Jiri Hron, Mark Rowland, Richard E. Turner, and Zoubin Ghahramani · 2018
Cited alongside, same era.
Moleculenet: a benchmark for molecular machine learning
Zhenqin Wu, Bharath Ramsundar, Evan N Feinberg, Joseph Gomes, Caleb Geniesse, Aneesh S Pappu, Karl Leswing, and Vijay Pande · 2018
Cited alongside, same era.
How powerful are graph neural networks?
Keyulu Xu, Weihua Hu, Jure Leskovec, and Stefanie Jegelka · 2018
Cited alongside, same era.
On exact computation with an infinitely wide neural net
Sanjeev Arora, Simon S Du, Wei Hu, Zhiyuan Li, Russ R Salakhutdinov, and Ruosong Wang · 2019
Cited alongside, same era.
Autoaugment: Learning augmentation strategies from data
Ekin D Cubuk, Barret Zoph, Dandelion Mane, Vijay Vasudevan, and Quoc V Le · 2019
Cited alongside, same era.
Finite versus infinite neural networks: an empirical study
Jaehoon Lee, Samuel Schoenholz, Jeffrey Pennington, Ben Adlam, Lechao Xiao, Roman Novak, and Jascha Sohl-Dickstein · 2020
Later among the works it cites.
Kernel methods through the roof: handling billions of points efficiently
Giacomo Meanti, Luigi Carratino, Lorenzo Rosasco, and Alessandro Rudi · 2020
Later among the works it cites.
Neural tangents: Fast and easy infinite neural networks in python
Roman Novak, Lechao Xiao, Jiri Hron, Jaehoon Lee, Alexander A. Alemi, Jascha Sohl-Dickstein, and Samuel S. Schoenholz · 2020
Later among the works it cites.
Neural kernels without tangents
Vaishaal Shankar, Alex Fang, Wenshuo Guo, Sara Fridovich-Keil, Jonathan Ragan-Kelley, Ludwig Schmidt, and Benjamin Recht · 2020
Later among the works it cites.
Tensor programs ii: Neural tangent kernel for any architecture
Greg Yang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Simon S Du, Kangcheng Hou, Russ R Salakhutdinov, Barnabas Poczos, Ruosong Wang, and Keyulu Xu · 2019
Cited alongside, same era.
Deep convolutional networks as shallow gaussian processes
Adrià Garriga-Alonso, Laurence Aitchison, and Carl Edward Rasmussen · 2019
Cited alongside, same era.
Enhanced convolutional neural tangent kernels
Zhiyuan Li, Ruosong Wang, Dingli Yu, Simon S Du, Wei Hu, Ruslan Salakhutdinov, and Sanjeev Arora · 2019
Cited alongside, same era.
Kernel machines that adapt to gpus for effective large batch training
Siyuan Ma and Mikhail Belkin · 2019
Cited alongside, same era.
Bayesian deep convolutional networks with many channels are gaussian processes
Roman Novak, Lechao Xiao, Jaehoon Lee, Yasaman Bahri, Greg Yang, Jiri Hron, Daniel A. Abolafia, Jeffrey Pennington, and Jascha Sohl-Dickstein · 2019
Cited alongside, same era.
A constructive prediction of the generalization error across scales
Jonathan S Rosenfeld, Amir Rosenfeld, Yonatan Belinkov, and Nir Shavit · 2019
Cited alongside, same era.
Efficientnet: Rethinking model scaling for convolutional neural networks
Mingxing Tan and Quoc Le · 2019
Cited alongside, same era.
Scaling neural tangent kernels via sketching and random features
Amir Zandieh, Insu Han, Haim Avron, Neta Shoham, Chaewon Kim, and Jinwoo Shin · 2020
Later among the works it cites.
Explaining neural scaling laws
Yasaman Bahri, Ethan Dyer, Jared Kaplan, Jaehoon Lee, and Utkarsh Sharma · 2021
Later among the works it cites.
Deep diversification of an aav capsid protein by machine learning
Drew Bryant, Ali Bashir, Sam Sinai, Nina Jain, Pierce Ogden, Patrick Riley, George Church, Lucy Colwell, and Eric Kelsic · 2021
Later among the works it cites.
FLIP: Benchmark tasks in fitness landscape inference for proteins
Christian Dallago, Jody Mou, Kadina E Johnston, Bruce Wittmann, Nick Bhattacharya, Samuel Goldman, Ali Madani, and Kevin K Yang · 2021
Later among the works it cites.
Reliable graph neural networks for drug discovery under distributional shift
Kehang Han, Balaji Lakshminarayanan, and Jeremiah Liu · 2021
Later among the works it cites.
Ogb-lsc: A large-scale challenge for machine learning on graphs
Weihua Hu, Matthias Fey, Hongyu Ren, Maho Nakata, Yuxiao Dong, and Jure Leskovec · 2021
Later among the works it cites.
The deep bootstrap framework: Good online learners are good offline generalizers
Preetum Nakkiran, Behnam Neyshabur, and Hanie Sedghi · 2021
Later among the works it cites.
Scaling laws for deep learning
Jonathan S Rosenfeld · 2021
Later among the works it cites.
Kernel interpolation for scalable online gaussian processes
Samuel Stanton, Wesley Maddox, Ian Delbridge, and Andrew Gordon Wilson · 2021
Later among the works it cites.
Efficientnetv2: Smaller models and faster training
Mingxing Tan and Quoc Le · 2021
Later among the works it cites.
Pooling architecture search for graph classification
Lanning Wei, Huan Zhao, Quanming Yao, and Zhiqiang He · 2021
Later among the works it cites.
Fast neural kernel embeddings for general activations
Insu Han, Amir Zandieh, Jaehoon Lee, Roman Novak, Lechao Xiao, and Amin Karbasi · 2022
Later among the works it cites.
Convolutional xformers for vision
Pranav Jeevan et al · 2022
Later among the works it cites.
Low-precision arithmetic for fast gaussian processes
Wesley J. Maddox, Andres Potapcynski, and Andrew Gordon Wilson · 2022
Later among the works it cites.