Fetching the paper…
Reading the bibliography…
Inspired by the Kolmogorov-Arnold representation theorem, we propose Kolmogorov-Arnold Networks (KANs) as promising alternatives to Multi-Layer Perceptrons (MLPs).
On the representation of continuous functions of several variables as superpositions of continuous functions of a smaller number of variables
A.N. Kolmogorov · 1956
Earlier work this paper cites.
On the representation of continuous functions of many variables by superposition of continuous functions of one variable and addition
Andrei Nikolaevich Kolmogorov · 1957
Earlier work this paper cites.
Absence of diffusion in certain random lattices
Philip W Anderson · 1958
Earlier work this paper cites.
A relation between the density of states and range of localization for one dimensional random systems
David J Thouless · 1972
Earlier work this paper cites.
A practical guide to splines
Carl De Boor · 1978
Earlier work this paper cites.
Scaling theory of localization: Absence of quantum diffusion in two dimensions
Elihu Abrahams, PW Anderson, DC Licciardello, and TV Ramakrishnan · 1979
Earlier work this paper cites.
Strong localization of photons in certain disordered dielectric superlattices
Sajeev John · 1987
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
George Cybenko · 1989
Earlier work this paper cites.
Multilayer feedforward networks are universal approximators
Kurt Hornik, Maxwell Stinchcombe, and Halbert White · 1989
Earlier work this paper cites.
Representation properties of networks: Kolmogorov’s theorem is irrelevant
Federico Girosi and Tomaso Poggio · 1989
Earlier work this paper cites.
Optimal nonlinear approximation
Ronald A DeVore, Ralph Howard, and Charles Micchelli · 1989
Earlier work this paper cites.
On the realization of a kolmogorov network
Ji-Nan Lin and Rolf Unbehauen · 1993
Earlier work this paper cites.
Wavelet compression and nonlinear n-widths
Ronald A DeVore, George Kyriazis, Dany Leviatan, and Vladimir M Tikhomirov · 1993
Earlier work this paper cites.
Neural networks: a comprehensive foundation
Simon Haykin · 1994
Earlier work this paper cites.
Brain plasticity and behavior
Bryan Kolb and Ian Q Whishaw · 1998
Earlier work this paper cites.
Space-filling curves and kolmogorov superposition-based neural networks
David A Sprecher and Sorin Draghici · 2002
Earlier work this paper cites.
On the training of a kolmogorov network
Mario Köppen · 2002
Earlier work this paper cites.
Riemannian Geometry
P. Petersen · 2006
Earlier work this paper cites.
Rate-optimal estimation for a general class of nonparametric regression models with unknown link functions
Joel L Horowitz and Enno Mammen · 2007
Earlier work this paper cites.
On a constructive proof of kolmogorov’s superposition theorem
Jürgen Braun and Michael Griebel · 2009
Earlier work this paper cites.
Fifty years of anderson localization
Ad Lagendijk, Bart van Tiggelen, and Diederik S Wiersma · 2009
Earlier work this paper cites.
Observation of a localization transition in quasiperiodic photonic lattices
Yoav Lahini, Rami Pugatch, Francesca Pozzi, Marc Sorel, Roberto Morandotti, Nir Davidson, and Yaron Silberberg · 2009
Earlier work this paper cites.
Modular and hierarchically modular organization of brain networks
David Meunier, Renaud Lambiotte, and Edward T Bullmore · 2010
Earlier work this paper cites.
Predicted mobility edges in one-dimensional incommensurate optical lattices: An exactly solvable model of anderson localization
J Biddle and S Das Sarma · 2010
Earlier work this paper cites.
Eureqa: software review
Renáta Dubcáková · 2011
Earlier work this paper cites.
The kolmogorov spline network for image processing
Pierre-Emmanuel Leni, Yohan D Fougerolle, and Frédéric Truchetet · 2013
Earlier work this paper cites.
Anderson localization of light
Mordechai Segev, Yaron Silberberg, and Demetrios N Christodoulides · 2013
Earlier work this paper cites.
Optics of photonic quasicrystals
Z Valy Vardeny, Ajay Nahata, and Amit Agrawal · 2013
Earlier work this paper cites.
Nonlinear material design using principal stretches
Hongyi Xu, Funshing Sin, Yufeng Zhu, and Jernej Barbič · 2015
Earlier work this paper cites.
Many-body localization and quantum nonergodicity in a model with a single-particle mobility edge
Xiaopeng Li, Sriram Ganeshan, JH Pixley, and S Das Sarma · 2015
Earlier work this paper cites.
Nearest neighbor tight binding models with an exact mobility edge in one dimension
Sriram Ganeshan, JH Pixley, and S Das Sarma · 2015
Earlier work this paper cites.
Absence of many-body mobility edges
Wojciech De Roeck, Francois Huveneers, Markus Müller, and Mauro Schiulaz · 2016
Earlier work this paper cites.
Extrapolation and learning equations
Georg Martius and Christoph H Lampert · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Why does deep and cheap learning work so well?
Henry W Lin, Max Tegmark, and David Rolnick · 2017
Earlier work this paper cites.
Error bounds for approximations with deep relu networks
Dmitry Yarotsky · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al · 2017
Earlier work this paper cites.
Deep sets
Manzil Zaheer, Satwik Kottur, Siamak Ravanbakhsh, Barnabas Poczos, Russ R Salakhutdinov, and Alexander J Smola · 2017
Earlier work this paper cites.
Deep learning scaling is predictable, empirically
Joel Hestness, Sharan Narang, Newsha Ardalani, Gregory Diamos, Heewoo Jun, Hassan Kianinejad, Md Mostofa Ali Patwary, Yang Yang, and Yanqi Zhou · 2017
Earlier work this paper cites.
Searching for activation functions
Prajit Ramachandran, Barret Zoph, and Quoc V Le · 2017
Earlier work this paper cites.
Algebraic multigrid methods
Jinchao Xu and Ludmil Zikatanov · 2017
Cited alongside, same era.
Relu deep neural networks and linear finite elements
Juncai He, Lin Li, Jinchao Xu, and Chunyue Zheng · 2018
Cited alongside, same era.
Measuring catastrophic forgetting in neural networks
Ronald Kemker, Marc McClure, Angelina Abitino, Tyler Hayes, and Christopher Kanan · 2018
Cited alongside, same era.
Optimizing kernel machines using deep learning
Huan Song, Jayaraman J Thiagarajan, Prasanna Sattigeri, and Andreas Spanias · 2018
Cited alongside, same era.
The deep ritz method: a deep learning-based numerical algorithm for solving variational problems
Bing Yu et al · 2018
Cited alongside, same era.
Nearly-tight vc-dimension and pseudodimension bounds for piecewise linear neural networks
Learning to Unknot
Sergei Gukov, James Halverson, Fabian Ruehle, and Piotr Sułkowski · 2021
Later among the works it cites.
Disentangling a deep learned volume formula
Jessica Craven, Vishnu Jejjala, and Arjun Kar · 2021
Later among the works it cites.
Multiscale invertible generative networks for high-dimensional bayesian inference
Shumao Zhang, Pengchuan Zhang, and Thomas Y Hou · 2021
Later among the works it cites.
Exsplinet: An interpretable and expressive spline-based neural network
Daniele Fakhoury, Emanuele Fakhoury, and Hendrik Speleers · 2022
Later among the works it cites.
How deep sparse networks avoid the curse of dimensionality: Efficiently computable functions are compositionally sparse
Tomaso Poggio · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Peter L Bartlett, Nick Harvey, Christopher Liaw, and Abbas Mehrabian · 2019
Cited alongside, same era.
Physics-informed neural networks: A deep learning framework for solving forward and inverse problems involving nonlinear partial differential equations
Maziar Raissi, Paris Perdikaris, and George E Karniadakis · 2019
Cited alongside, same era.
Learning activation functions: A new paradigm for understanding neural networks
Mohit Goyal, Rajan Goyal, and Brejesh Lall · 2019
Cited alongside, same era.
Deep spline networks with control of lipschitz regularity
Shayan Aziznejad and Michael Unser · 2019
Cited alongside, same era.
Error bounds for deep relu networks using the kolmogorov–arnold superposition theorem
Hadrien Montanelli and Haizhao Yang · 2020
Cited alongside, same era.
Theoretical issues in deep networks
Tomaso Poggio, Andrzej Banburski, and Qianli Liao · 2020
Cited alongside, same era.
A neural scaling law from the dimension of the data manifold
Utkarsh Sharma and Jared Kaplan · 2020
Cited alongside, same era.
Catherine Olsson, Nelson Elhage, Neel Nanda, Nicholas Joseph, Nova DasSarma, Tom Henighan, Ben Mann, Amanda Askell, Yuntao Bai, Anna Chen, et al · 2022
Later among the works it cites.
Locating and editing factual associations in gpt
Kevin Meng, David Bau, Alex Andonian, and Yonatan Belinkov · 2022
Later among the works it cites.
Nelson Elhage, Tristan Hume, Catherine Olsson, Nicholas Schiefer, Tom Henighan, Shauna Kravec, Zac Hatfield-Dodds, Robert Lasenby, Dawn Drain, Carol Chen, et al · 2022
Later among the works it cites.
Softmax linear units
Nelson Elhage, Tristan Hume, Catherine Olsson, Neel Nanda, Tom Henighan, Scott Johnston, Sheer ElShowk, Nicholas Joseph, Nova DasSarma, Ben Mann, Danny Hernandez, Amanda Askell, Kamal Ndousse, Andy Jones, Dawn Drain, Anna Chen, Yuntao Bai, Deep Ganguli, Liane Lovitt, Zac Hatfield-Dodds, Jackson Kernion, Tom Conerly, Shauna Kravec, Stanislav Fort, Saurav Kadavath, Josh Jacobson, Eli Tran-Johnson, Jared Kaplan, Jack Clark, Tom Brown, Sam McCandlish, Dario Amodei, and Christopher Olah · 2022
Later among the works it cites.
Neural network architecture beyond width and depth
Shijun Zhang, Zuowei Shen, and Haizhao Yang · 2022
Later among the works it cites.
Discovering parametric activation functions
Garrett Bingham and Risto Miikkulainen · 2022
Later among the works it cites.
Fourier continuation for exact derivative computation in physics-informed neural operators
Haydn Maust, Zongyi Li, Yixuan Wang, Daniel Leibovici, Oscar Bruno, Thomas Hou, and Anima Anandkumar · 2022
Later among the works it cites.
Illuminating new and known relations between knot invariants
Jessica Craven, Mark Hughes, Vishnu Jejjala, and Arjun Kar · 2022
Later among the works it cites.
Sparse autoencoders find highly interpretable features in language models
Hoagy Cunningham, Aidan Ewart, Logan Riggs, Robert Huben, and Lee Sharkey · 2023
Later among the works it cites.
Juncai He · 2023
Later among the works it cites.
Deep neural networks and finite elements of any order on arbitrary dimensions
Juncai He and Jinchao Xu · 2023
Later among the works it cites.
Precision machine learning
Eric J Michaud, Ziming Liu, and Max Tegmark · 2023
Later among the works it cites.
Optimal approximation rates for deep relu neural networks on sobolev and besov spaces
Jonathan W Siegel · 2023
Later among the works it cites.
Searching for ribbons with machine learning, 2023
Sergei Gukov, James Halverson, Ciprian Manolescu, and Fabian Ruehle · 2023
Later among the works it cites.
Reentrant delocalization transition in one-dimensional photonic quasicrystals
Sachin Vaidya, Christina Jörg, Kyle Linn, Megan Goh, and Mikael C Rechtsman · 2023
Later among the works it cites.
Exact new mobility edges between critical and localized states
Xin-Chi Zhou, Yongjian Wang, Ting-Fung Jeffrey Poon, Qi Zhou, and Xiong-Jun Liu · 2023
Later among the works it cites.
A new iterative method for construction of the kolmogorov-arnold representation
Michael Poluektov and Andrew Polar · 2023
Later among the works it cites.
The quantization model of neural scaling
Eric J Michaud, Ziming Liu, Uzay Girit, and Max Tegmark · 2023
Later among the works it cites.
Interpretability in the wild: a circuit for indirect object identification in GPT-2 small
Kevin Ro Wang, Alexandre Variengien, Arthur Conmy, Buck Shlegeris, and Jacob Steinhardt · 2023
Later among the works it cites.
Progress measures for grokking via mechanistic interpretability
Neel Nanda, Lawrence Chan, Tom Lieberum, Jess Smith, and Jacob Steinhardt · 2023
Later among the works it cites.
The clock and the pizza: Two stories in mechanistic explanation of neural networks
Ziqian Zhong, Ziming Liu, Max Tegmark, and Jacob Andreas · 2023
Later among the works it cites.
Seeing is believing: Brain-inspired modular training for mechanistic interpretability
Ziming Liu, Eric Gan, and Max Tegmark · 2023
Later among the works it cites.
Interpretable machine learning for science with pysr and symbolicregression. jl
Miles Cranmer · 2023
Later among the works it cites.
Neural operator: Learning maps between function spaces with applications to pdes
Nikola Kovachki, Zongyi Li, Burigede Liu, Kamyar Azizzadenesheli, Kaushik Bhattacharya, Andrew Stuart, and Anima Anandkumar · 2023
Later among the works it cites.
Machine Learning in Pure Mathematics and Theoretical Physics
Y.H. He · 2023
Later among the works it cites.
Exponentially convergent multiscale finite element method
Yifan Chen, Thomas Y Hou, and Yixuan Wang · 2023
Later among the works it cites.
Sharp lower bounds on the manifold widths of sobolev and besov spaces
Jonathan W Siegel · 2024
Closest in time.
Multi-stage neural networks: Function approximator of machine precision
Yongji Wang and Ching-Yao Lai · 2024
Closest in time.
Revisiting neural networks for continual learning: An architectural perspective, 2024
Aojun Lu, Tao Feng, Hangjie Yuan, Xiaotian Song, and Yanan Sun · 2024
Closest in time.
On the kolmogorov neural networks
Aysu Ismayilova and Vugar E Ismailov · 2024
Closest in time.
A resource model for neural scaling law
Jinyeop Song, Ziming Liu, Max Tegmark, and Jeff Gore · 2024
Closest in time.
https://github.com/trevorstephens/gplearn
Gplearn · 2024
Closest in time.
Separable physics-informed neural networks
Junwoo Cho, Seungtae Nam, Hyunmo Yang, Seok-Bae Yun, Youngjoon Hong, and Eunbyung Park · 2024
Closest in time.
Rigor with machine learning from field theory to the poincaréconjecture
Sergei Gukov, James Halverson, and Fabian Ruehle · 2024
Closest in time.