Fetching the paper…
Reading the bibliography…
How can we build agents that keep learning from experience, quickly and efficiently, after their initial training? Here we take inspiration from the main mechanism of learning in biological brains: synaptic plasticity, carefully tuned by evolution to produce efficient lifelong learning.
The organization of behavior: a neuropsychological theory
Hebb, D. O · 1949
Earlier work this paper cites.
Neural networks and physical systems with emergent collective computational abilities
Hopfield, J. J · 1982
Earlier work this paper cites.
Learning a synaptic learning rule
Bengio, Y., Bengio, S., and Cloutier, J · 1991
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Learning to learn
Thrun, S. and Pratt, L · 1998
Earlier work this paper cites.
Synaptic plasticity and memory: an evaluation of the hypothesis
Martin, S. J., Grimwood, P. D., and Morris, R. G · 2000
Earlier work this paper cites.
Theoretical neuroscience , volume 806
Dayan, P. and Abbott, L. F · 2001
Earlier work this paper cites.
Learning to learn using gradient descent
Hochreiter, S., Younger, A., and Conwell, P · 2001
Earlier work this paper cites.
By carrot or by stick: cognitive reinforcement learning in parkinsonism
Frank, M. J., Seeberger, L. C., and O’reilly, R. C · 2004
Earlier work this paper cites.
Oja learning rule
Oja, E · 2008
Earlier work this paper cites.
Evolutionary advantages of neuromodulated plasticity in dynamic, reward-based scenarios
Soltoggio, A., Bullinaria, J. A., Mattiussi, C., Dürr, P., and Floreano, D · 2008
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Cited alongside, same era.
Optogenetic stimulation of a hippocampal engram activates fear memory recall
Liu, X., Ramirez, S., Pang, P. T., Puryear, C. B., Govindarajan, A., Deisseroth, K., and Tonegawa, S · 2012
Cited alongside, same era.
Neural turing machines
Graves, A., Wayne, G., and Danihelka, I · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2015
Cited alongside, same era.
Human-level concept learning through probabilistic program induction
Lake, B. M., Salakhutdinov, R., and Tenenbaum, J. B · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al · 2015
One-shot learning with Memory-Augmented neural networks
Santoro, A., Bartunov, S., Botvinick, M., Wierstra, D., and Lillicrap, T · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al · 2016
Later among the works it cites.
Matching networks for one shot learning
Vinyals, O., Blundell, C., Lillicrap, T., Wierstra, D., et al · 2016
Later among the works it cites.
Learning to reinforcement learn
Wang, J. X., Kurth-Nelson, Z., Tirumala, D., Soyer, H., Leibo, J. Z., Munos, R., Blundell, C., Kumaran, D., and Botvinick, M · 2016
Later among the works it cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S · 2017
Later among the works it cites.
Learning to remember rare events
Kaiser, L., Nachum, O., Roy, A., and Bengio, S · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
End-To-End memory networks
Sukhbaatar, S., Szlam, A., Weston, J., and Fergus, R · 2015
Cited alongside, same era.
Using fast weights to attend to the recent past
Ba, J., Hinton, G. E., Mnih, V., Leibo, J. Z., and Ionescu, C · 2016
Cited alongside, same era.
Rl 2 : Fast reinforcement learning via slow reinforcement learning
Duan, Y., Schulman, J., Chen, X., Bartlett, P. L., Sutskever, I., and Abbeel, P · 2016
Cited alongside, same era.
Backpropagation of hebbian plasticity for continual learning
Miconi, T · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
Mnih, V., Badia, A. P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K · 2016
Cited alongside, same era.
A ‘self-referential’weight matrix
Schmidhuber, J
Cited in the paper.
Later among the works it cites.
A simple neural attentive meta-learner
Mishra, N., Rohaninejad, M., Chen, X., and Abbeel, P · 2017
Later among the works it cites.
Gated fast weights for on-the-fly neural program generation
Schlag, I. and Schmidhuber, J · 2017
Later among the works it cites.
Prototypical networks for few-shot learning
Snell, J., Swersky, K., and Zemel, R · 2017
Later among the works it cites.
Born to learn: the inspiration, progress, and future of evolved plastic artificial neural networks
Soltoggio, A., Stanley, K. O., and Risi, S · 2017
Later among the works it cites.
Reducing the ratio between learning complexity and number of time varying variables in fully recurrent nets
Schmidhuber, J · 2063
Closest in time.