Fetching the paper…
Reading the bibliography…
Inner products of neural network feature maps arise in a wide variety of machine learning frameworks as a method of modeling relations between inputs.
“Functions of Positive and Negative Type, and Their Connection with the Theory of Integral Equations”
J. Mercer · 1909
Earlier work this paper cites.
“Non-Symmetric Kernels of Positive Type”
Caroline. Seely · 1919
Earlier work this paper cites.
“Theory of reproducing kernels”
Nachman Aronszajn · 1950
Earlier work this paper cites.
“Representation of a Preference Ordering by a Numerical Function”
Gerard Debreu · 1954
Earlier work this paper cites.
“Existence of a Continuous Utility Function: An Elementary Proof”
Jean-Yves Jaffray · 1975
Earlier work this paper cites.
“Learning Representations by Back-Propagating Errors”
David. Rumelhart, Geoffrey. Hinton and Ronald. Williams · 1986
Earlier work this paper cites.
“A Time-Delay Neural Network Architecture for Speech Recognition”
Kevin. Lang · 1988
Earlier work this paper cites.
“Approximation by Superpositions of a Sigmoidal Function”
G. Cybenko · 1989
Earlier work this paper cites.
“Neural Networks for Fingerprint Recognition”
Pierre Baldi and Yves Chauvin · 1993
Earlier work this paper cites.
“Universal Approximation Bounds for Superpositions of a Sigmoidal Function”
A.R. Barron · 1993
Earlier work this paper cites.
“Signature Verification Using a” Siamese” Time Delay Neural Network”
Jane Bromley, Isabelle Guyon, Yann LeCun, Eduard S“”ackinger and Roopak Shah · 1993
Earlier work this paper cites.
“Nonlinear approximation”
Ronald DeVore · 1998
Earlier work this paper cites.
“Uniform approximation by neural networks”
Y. Makovoz · 1998
Earlier work this paper cites.
“Approximation by ridge functions and neural networks”
P.. Petrushev · 1998
Earlier work this paper cites.
“Lower bounds for approximation by MLP neural networks”
Vitaly Maiorov and Allan Pinkus · 1999
Earlier work this paper cites.
“Approximation theory of the MLP model in neural networks”
A. Pinkus · 1999
Earlier work this paper cites.
“On the near optimality of the stochastic approximation of smooth functions by neural networks”
VE Maiorov and Ron Meir · 2000
Earlier work this paper cites.
“Error bounds for approximation with neural networks”
M. Burger and A. Neubauer · 2001
Cited alongside, same era.
“Mercer Theorem for RKHS on Noncompact Sets”
Hongwei Sun · 2004
Cited alongside, same era.
“Learning a similarity metric discriminatively, with application to face verification”
S. Chopra, R. Hadsell and Y. LeCun · 2005
Cited alongside, same era.
“Object-Centric Learning with Slot Attention”
Francesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran, Georg Heigold, Jakob Uszkoreit, Alexey Dosovitskiy and Thomas Kipf · 2006
Cited alongside, same era.
“Approximation by neural networks and learning theory”
V. Maiorov · 2006
Cited alongside, same era.
“Universal Kernels”
Charles. Micchelli, Yuesheng Xu and Haizhang Zhang · 2006
Cited alongside, same era.
“Bert: Pre-training of Deep Bidirectional Transformers for Language Understanding”, 2018
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2018
Later among the works it cites.
“Speech-Transformer: A No-Recurrence Sequence-to-Sequence Model for Speech Recognition”
Linhao Dong, Shuang Xu and Bo Xu · 2018
Later among the works it cites.
“Relational Recurrent Neural Networks”
Adam Santoro et al · 2018
Later among the works it cites.
“An Explicit Neural Network Construction for Piecewise Constant Function Approximation”
Kailiang Wu and Dongbin Xiu · 2018
Later among the works it cites.
“Deep Reinforcement Learning with Relational Inductive Biases”
Vinicius Zambaldi et al · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Reproducing Kernel Banach Spaces for Machine Learning”
Haizhang Zhang, Yuesheng Xu and Jun Zhang · 2009
Cited alongside, same era.
“An Image Is Worth 16x16 Words: Transformers for Image Recognition at Scale”, 2020
Alexey Dosovitskiy et al · 2010
Cited alongside, same era.
“Emergent Symbols through Binding in External Memory”, 2021
Taylor. Webb, Ishan Sinha and Jonathan. Cohen · 2012
Cited alongside, same era.
“Neural Turing Machines”
Alex Graves, Greg Wayne and Ivo Danihelka · 2014
Cited alongside, same era.
“Siamese Neural Networks for One-Shot Image Recognition”
Gregory Koch, Richard Zemel and Ruslan Salakhutdinov · 2015
Cited alongside, same era.
“Breaking the Curse of Dimensionality with Convex Neural Networks”
Francis Bach · 2016
Cited alongside, same era.
“Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer”
Colin Raffel et al · 2020
Later among the works it cites.
“Swin Transformer: Hierarchical Vision Transformer Using Shifted Windows”
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin and Baining Guo · 2021
Later among the works it cites.
“Transformers are Deep Infinite-Dimensional Non-Mercer Binary Kernel Machines”, 2021
Matthew. Wright and Joseph. Gonzalez · 2021
Later among the works it cites.
“Scaling Instruction-Finetuned Language Models”
Hyung Chung et al · 2022
Later among the works it cites.
“On Neural Architecture Inductive Biases for Relational Tasks”, 2022
Giancarlo Kerg, Sarthak Mittal, David Rolnick, Yoshua Bengio, Blake Richards and Guillaume Lajoie · 2022
Later among the works it cites.
OpenAI · 2023
Later among the works it cites.
“Alpaca: A Strong, Replicable Instruction-Following Model” Stanford Center for Research on Foundation Models, 2023
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang and Tatsunori. Hashimoto · 2023
Later among the works it cites.
“Llama 2: Open Foundation and Fine-Tuned Chat Models”
Hugo Touvron et al · 2023
Later among the works it cites.
Awni Altabaa and John Lafferty · 2024
Closest in time.
“Learning Hierarchical Relational Representations through Relational Convolutions”, 2024
Awni Altabaa and John Lafferty · 2024
Closest in time.
“Abstractors and relational cross-attention: An inductive bias for explicit relational reasoning in Transformers”
Awni Altabaa, Taylor Webb, Jonathan Cohen and John Lafferty · 2024
Closest in time.