Fetching the paper…
Reading the bibliography…
We introduce a modern Hopfield network with continuous states and a corresponding update rule.
Interpreting and improving natural-language processing (in machines) with natural language-processing (in the brain)
M. Toneva and L. Wehbe · 1905
Earlier work this paper cites.
Nonlinear programming: a unified approach
W. I. Zangwill · 1969
Earlier work this paper cites.
Sufficient conditions for the convergence of monotonic mathematical programming algorithms
R. R. Meyer · 1976
Earlier work this paper cites.
Analytic theory of the ground state properties of a spin glass. I. Ising spin glass
F. Tanaka and S. F. Edwards · 1980
Earlier work this paper cites.
Neural networks and physical systems with emergent collective computational abilities
J. J. Hopfield · 1982
Earlier work this paper cites.
On the convergence properties of the em algorithm
J. C. F. Wu · 1983
Earlier work this paper cites.
Neurons with graded response have collective computational properties like those of two-state neurons
J. J. Hopfield · 1984
Earlier work this paper cites.
Information capacity of the Hopfield model
Y. Abu-Mostafa and J.-M. StJacques · 1985
Earlier work this paper cites.
Saturation level of the Hopfield model for neural network
A. Crisanti, D. J. Amit, and H. Gutfreund · 1986
Earlier work this paper cites.
The capacity of the Hopfield associative memory
R. J. McEliece, E. C. Posner, E. R. Rodemich, and S. S. Venkatesh · 1987
Earlier work this paper cites.
On the number of spurious memories in the Hopfield model
J. Bruck and V. P. Roychowdhury · 1990
Earlier work this paper cites.
Introduction to the Theory of Neural Computation
J. Hertz, A. Krogh, and R. G. Palmer · 1991
Earlier work this paper cites.
Untersuchungen zu dynamischen neuronalen Netzen. Diploma thesis, Institut für Informatik, Lehrstuhl Prof. Brauer, Technische Universität München, 1991
S. Hochreiter · 1991
Earlier work this paper cites.
Learning to control fast-weight memories: An alternative to dynamic recurrent networks
J. Schmidhuber · 1992
Earlier work this paper cites.
Mining association rules between sets of items in large databases
R. Agrawal, T. Imieliundefinedski, and A. Swami · 1993
Earlier work this paper cites.
Dynamics of discrete time, continuous state Hopfield networks
P. Koiran · 1994
Earlier work this paper cites.
Support-vector networks
C. Cortes and V. Vapnik · 1995
Earlier work this paper cites.
A novel optimizing network architecture with applications
A. Rangarajan, S. Gold, and E. Mjolsness · 1996
Earlier work this paper cites.
Solving the multiple instance problem with axis-parallel rectangles
T. G. Dietterich, R. H. Lathrop, and T. Lozano-Pérez · 1997
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
On the storage capacity of nonlinear neural networks
C. Mazza · 1997
Earlier work this paper cites.
A framework for multiple-instance learning
O. Maron and T. Lozano-Pérez · 1998
Earlier work this paper cites.
Convergence properties of the softassign quadratic assignment algorithm
A. Rangarajan, A. Yuille, and Eric E. Mjolsness · 1999
Earlier work this paper cites.
Solving the multiple-instance problem: A lazy learning approach
J. Wang · 2000
Earlier work this paper cites.
Random forests
L. Breiman · 2001
Earlier work this paper cites.
Learning with Kernels – Support Vector Machines, Regularization, Optimization, and Beyond
B. Schölkopf and A. J. Smola · 2002
Earlier work this paper cites.
Storage capacity of attractor neural networks with depressing synapses
J. J. Torres, L. Pantic, and Hilbert H. J. Kappen · 2002
Earlier work this paper cites.
The concave-convex procedure (CCCP)
A. L. Yuille and A. Rangarajan · 2002
Earlier work this paper cites.
Support vector machines for multiple-instance learning
S. Andrews, I. Tsochantaridis, and T. Hofmann · 2003
Earlier work this paper cites.
An isotropic Gaussian mixture can have more modes than components
M. Carreira-Perpiñán and C. K. I. Williams · 2003
Earlier work this paper cites.
The concave-convex procedure
A. L. Yuille and A. Rangarajan · 2003
Earlier work this paper cites.
MILES: Multiple-instance learning via embedded instance selection
Y. Chen, J. Bi, and J. Z. Wang · 2006
Earlier work this paper cites.
Modern Hopfield networks and attention for immune repertoire classification
M. Widrich, B. Schäfl, M. Pavlović, H. Ramsauer, L. Gruber, M. Holzleitner, J. Brandstetter, G. K. Sandve, V. Greiff, S. Hochreiter, and G. Klambauer · 2007
Earlier work this paper cites.
Inequalities on the Lambert w w function and hyperpower function
A. Hoorfar and M. Hassani · 2008
Earlier work this paper cites.
Convex Optimization
S. Boyd and L. Vandenberghe · 2009
Earlier work this paper cites.
On the convergence of the concave-convex procedure
B. K. Sriperumbudur and G. R. Lanckriet · 2009
Earlier work this paper cites.
NIST handbook of mathematical functions
F. W. J. Olver, D. W. Lozier, R. F. Boisvert, and C. W. Clark · 2010
Cited alongside, same era.
A Bayesian approach to in silico blood-brain barrier penetration modeling
I. F. Martins, A. L. Teixeira, L. Pinheiro, and A. O. Falcao · 2012
Cited alongside, same era.
Distributions of angles in random packing on spheres
T. Cai, J. Fan, and T. Jiang · 2013
Cited alongside, same era.
Topological and dynamical complexity of random neural networks
G. Wainrib and J. Touboul · 2013
Cited alongside, same era.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
K. Cho, B. vanMerriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Cited alongside, same era.
Do we need hundreds of classifiers to solve real world classification problems?
M. Fernández-Delgado, E. Cernadas, S. Barro, and D. Amorim · 2014
PointNet: Deep learning on point sets for 3d classification and segmentation
C. R. Qi, H. Su, M. Kaichun, and L. J. Guibas · 2017
Later among the works it cites.
MoleculeNet: A benchmark for molecular machine learning
Z. Wu, B. Ramsundar, E. N. Feinberg, J. Gomes, C. Geniesse, A. S. Pappu, K. Leswing, and V. Pande · 2017
Later among the works it cites.
Deep sets
M. Zaheer, S. Kottur, S. Ravanbakhsh, B. Poczos, R. R. Salakhutdinov, and A. J. Smola · 2017
Later among the works it cites.
Learning to update auto-associative memory in recurrent neural networks for improving sequence memorization
W. Zhang and B. Zhou · 2017
Later among the works it cites.
Sharp bounds for the lambert w w function
F. Alzahrani and A. Salem · 2018
Later among the works it cites.
A new mechanical approach to handle generalized Hopfield neural networks
A. Barra, M. Beccaria, and A. Fachechi · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Neural turing machines
A. Graves, G. Wayne, and I. Danihelka · 2014
Cited alongside, same era.
Empowering multiple instance histopathology cancer diagnosis by cell graphs
M. Kandemir, C. Zhang, and F. A. Hamprecht · 2014
Cited alongside, same era.
Deep learning in neural networks: An overview
J. Schmidhuber · 2014
Cited alongside, same era.
Memory networks
J. Weston, S. Chopra, and A. Bordes · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2015
Cited alongside, same era.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Cited alongside, same era.
Later among the works it cites.
Multiple instance learning: a survey of problem characteristics and applications
M.-A. Carbonneau, V. Cheplygina, E. Granger, and G. Gagnon · 2018
Later among the works it cites.
BERT: pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Later among the works it cites.
Attention-based deep multiple instance learning
M. Ilse, J. M. Tomczak, and M. Welling · 2018
Later among the works it cites.
Study and observation of the variation of accuracies of KNN, SVM, LMNN, ENN algorithms on eleven different datasets from UCI machine learning repository
M. M. R. Khan, R. B. Arif, M. A. B. Siddique, and M. R. Oishe · 2018
Later among the works it cites.
BRUNO: A deep recurrent model for exchangeable data
I. Korshunova, J. Degrave, F. Huszar, Y. Gal, A. Gretton, and J. Dambre · 2018
Later among the works it cites.
Dense associative memory is robust to adversarial inputs
D. Krotov and J. J. Hopfield · 2018
Later among the works it cites.
Bag encoding strategies in multiple instance learning problems
E. Ş. Küçükaşcı and M. G. Baydoğan · 2018
Later among the works it cites.
Learning to reason with third order tensor products
I. Schlag and J. Schmidhuber · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction
R. S. Sutton and A. G. Barto · 2018
Later among the works it cites.
Graph attention networks
P. Velic̆ković, G. Cucurull, A. Casanova, A. Romero, P. Liò, and Y. Bengio · 2018
Later among the works it cites.
Revisiting multiple instance neural networks
X. Wang, Y. Yan, P. Tang, X. Bai, and W. Liu · 2018
Later among the works it cites.
Improved expressivity through dendritic neural networks
X. Wu, X. Liu, W. Li, and Q. Wu · 2018
Later among the works it cites.
SpiderCNN: Deep learning on point sets with parameterized convolutional filters
Y. Xu, T. Fan, M. Xu, L. Zeng, and Y. Qiao · 2018
Later among the works it cites.
A compact vocabulary of paratope-epitope interactions enables predictability of antibody-antigen binding
R. Akbar, P. A. Robert, M. Pavlović, J. R. Jeliazkov, I. Snapkov, A. Slabodkin, C. R. Weber, L. Scheffer, E. Miho, I. H. Haff, et al · 2019
Later among the works it cites.
Universal transformers
M. Dehghani, S. Gouws, O. Vinyals, J. Uszkoreit, and L. Kaiser · 2019
Later among the works it cites.
BERT: pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Later among the works it cites.
PyTorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, et al · 2019
Later among the works it cites.
Enhancing the transformer with explicit relational encoding for math problem solving
I. Schlag, P. Smolensky, R. Fernandez, N. Jojic, J. Schmidhuber, and J. Gao · 2019
Later among the works it cites.
HuggingFace’s transformers: State-of-the-art natural language processing
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, and J. Brew · 2019
Later among the works it cites.
MEMO: a deep network for flexible combination of episodic memories
A. Banino, A. P. Badia, R. Köster, M. J. Chadwick, V. Zambaldi, D. Hassabis, C. Barry, M. Botvinick, D. Kumaran, and C. Blundell · 2020
Closest in time.
Encoding-based memory modules for recurrent neural networks
A. Carta, A. Sperduti, and D. Bacciu · 2020
Closest in time.
ELECTRA: Pre-training text encoders as discriminators rather than generators
K. Clark, M.-T. Luong, Q. V. Le, and C. D. Manning · 2020
Closest in time.
Deep multiple instance learning for digital histopathology
M. Ilse, J. M. Tomczak, and M. Welling · 2020
Closest in time.
Could graph neural networks learn better molecular representation for drug discovery? a comparison study of descriptor-based and graph-based models
D. Jiang, Z. Wu, C.-Y. Hsieh, G. Chen, B. Liao, Z. Wang, C. Shen, D. Cao, J. Wu, and T. Hou · 2020
Closest in time.
Large associative memory problem in neurobiology and machine learning
D. Krotov and J. J. Hopfield · 2020
Closest in time.
Synthesizer: Rethinking self-attention in transformer models
Y. Tay, D. Bahri, D. Metzler, D.-C. Juan, Z. Zhao, and C. Zheng · 2020
Closest in time.
immuneSIM: tunable multi-feature simulation of B- and T-cell receptor repertoires for immunoinformatics benchmarking
C. R. Weber, R. Akbar, A. Yermanos, M. Pavlović, I. Snapkov, G. K. Sandve, S. T. Reddy, and V. Greiff · 2020
Closest in time.
Pushing the boundaries of molecular representation for drug discovery with the graph attention mechanism
Z. Xiong, D. Wang, X. Liu, F. Zhong, X. Wan, X. Li, Z. Li, X. Luo, K. Chen, H. Jiang, and M. Zheng · 2020
Closest in time.
Set distribution networks: a generative model for sets of images
S. Zhai, W. Talbott, M. A. Bautista, C. Guestrin, and J. M. Susskind · 2020
Closest in time.
Linear transformers are secretly fast weight memory systems
I. Schlag, K. Irie, and J. Schmidhuber · 2021
Closest in time.