Fetching the paper…
Reading the bibliography…
Artificial Neural Networks are uniquely adroit at machine learning by processing data through a network of artificial neurons.
On computable numbers, with an application to the entscheidungsproblem
A.M Turing · 1936
Earlier work this paper cites.
The mindful brain: cortical organization and the group-selective theory of higher brain function
Gerald M Edelman and Vernon B Mountcastle · 1978
Earlier work this paper cites.
The modular operation of the cerebral neocortex considered as the material basis of mental events
JC Eccles · 1981
Earlier work this paper cites.
The correlation theory of brain function, 1981
Christoph von der Malsburg · 1981
Earlier work this paper cites.
Learning representations by back-propagating errors
David E Rumelhart, Geoffrey E Hinton, and Ronald J Williams · 1986
Earlier work this paper cites.
Eye, Brain, and Vision
D.H. Hubel · 1988
Earlier work this paper cites.
Making the world differentiable: On using self-supervised fully recurrent n eu al networks for dynamic reinforcement learning and planning in non-stationary environm nts
Jiirgen Schmidhuber · 1990
Earlier work this paper cites.
Adaptive mixtures of local experts
Robert A Jacobs, Michael I Jordan, Steven J Nowlan, and Geoffrey E Hinton · 1991
Earlier work this paper cites.
Learning to control fast-weight memories: An alternative to dynamic recurrent networks
Jürgen Schmidhuber · 1992
Earlier work this paper cites.
Neural darwinism: selection and reentrant signaling in higher brain function
Gerald M Edelman · 1993
Earlier work this paper cites.
First draft of a report on the edvac
John Von Neumann · 1993
Earlier work this paper cites.
Design and evolution of modular neural network architectures
Bart LM Happel and Jacob MJ Murre · 1994
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Cited alongside, same era.
Catastrophic forgetting in connectionist networks
Robert M French · 1999
Cited alongside, same era.
Human brain function
Richard SJ Frackowiak · 2004
Cited alongside, same era.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al · 2009
Cited alongside, same era.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Cited alongside, same era.
Estimating or propagating gradients through stochastic neurons for conditional computation
Yoshua Bengio, Nicholas Léonard, and Aaron Courville · 2013
Meta-learning with memory-augmented neural networks
Adam Santoro, Sergey Bartunov, Matthew Botvinick, Daan Wierstra, and Timothy Lillicrap · 2016
Later among the works it cites.
Hypernetworks
David Ha, Andrew M. Dai, and Quoc V. Le · 2017
Later among the works it cites.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q Weinberger · 2017
Later among the works it cites.
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer
Noam Shazeer, Azalia Mirhoseini, Krzysztof Maziarz, Andy Davis, Quoc V. Le, Geoffrey E. Hinton, and Jeff Dean · 2017
Later among the works it cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Cited alongside, same era.
Alex Graves, Greg Wayne, and Ivo Danihelka · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Cited alongside, same era.
Hybrid computing using a neural network with dynamic external memory
Alex Graves, Greg Wayne, Malcolm Reynolds, Tim Harley, Ivo Danihelka, Agnieszka Grabska-Barwińska, Sergio Gómez Colmenarejo, Edward Grefenstette, Tiago Ramalho, John Agapiou, et al · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Friedemann Zenke, Ben Poole, and Surya Ganguli · 2017
Later among the works it cites.
Re-evaluating continual learning scenarios: A categorization and case for strong baselines
Yen-Chang Hsu, Yen-Cheng Liu, Anita Ramasamy, and Zsolt Kira · 2018
Later among the works it cites.
Variational memory encoder-decoder
Hung Le, Truyen Tran, Thin Nguyen, and Svetha Venkatesh · 2018
Later among the works it cites.
Learning to remember more with less memorization
Hung Le, Truyen Tran, and Svetha Venkatesh · 2018
Later among the works it cites.
Routing networks: Adaptive selection of non-linear functions for multi-task learning
Clemens Rosenbaum, Tim Klinger, and Matthew Riemer · 2018
Later among the works it cites.
Neural stored-program memory
Hung Le, Truyen Tran, and Svetha Venkatesh · 2020
Closest in time.
Self-attentive associative memory
Hung Le, Truyen Tran, and Svetha Venkatesh · 2020
Closest in time.