Fetching the paper…
Reading the bibliography…
Continual Learning in artificial neural networks suffers from interference and forgetting when different tasks are learned sequentially.
Loss of recent memory after bilateral hippocampal lesions
William Beecher Scoville and Brenda Milner · 1957
Earlier work this paper cites.
Effects of visual deprivation on morphology and physiology of cells in the cat’s lateral geniculate body
Torsten N Wiesel, David H Hubel, et al · 1963
Earlier work this paper cites.
Society of mind
Marvin Minsky · 1988
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
Michael McCloskey and Neal J Cohen · 1989
Earlier work this paper cites.
Theory of the backpropagation neural network
Robert Hecht-Nielsen · 1989
Earlier work this paper cites.
Handwritten digit recognition with a back-propagation network
B Boser Le Cun, John S Denker, D Henderson, Richard E Howard, W Hubbard, and Lawrence D Jackel · 1990
Earlier work this paper cites.
Direct transfer of learned information among neural networks
Lorien Y Pratt, Jack Mostow, Candace A Kamm, and Ace A Kamm · 1991
Earlier work this paper cites.
Self-improving reactive agents based on reinforcement learning, planning and teaching
Long-Ji Lin · 1992
Earlier work this paper cites.
Signature verification using a “siamese” time delay neural network
Jane Bromley, James W Bentz, Léon Bottou, Isabelle Guyon, Yann LeCun, Cliff Moore, Eduard Säckinger, and Roopak Shah · 1993
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: insights from the successes and failures of connectionist models of learning and memory
James L McClelland, Bruce L McNaughton, and Randall C O’Reilly · 1995
Earlier work this paper cites.
Learning many related tasks at the same time with backpropagation
Rich Caruana · 1995
Earlier work this paper cites.
Born again trees
Leo Breiman and Nong Shang · 1996
Earlier work this paper cites.
Multitask learning
Rich Caruana · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Catastrophic forgetting in connectionist networks
Robert M French · 1999
Earlier work this paper cites.
Thermodynamic depth of causal states: Objective complexity via minimal representations
James P Crutchfield and Cosma Rohilla Shalizi · 1999
Earlier work this paper cites.
Causal architecture, complexity and self-organization in the time series and cellular automata
Cosma Rohilla Shalizi et al · 2001
Earlier work this paper cites.
Distributed and overlapping representations of faces and objects in ventral temporal cortex
James V Haxby, M Ida Gobbini, Maura L Furey, Alumit Ishai, Jennifer L Schouten, and Pietro Pietrini · 2001
Cited alongside, same era.
Gradient flow in recurrent nets: the difficulty of learning long-term dependencies, 2001
Sepp Hochreiter, Yoshua Bengio, Paolo Frasconi, and Jürgen Schmidhuber · 2001
Cited alongside, same era.
The organization of behavior: A neuropsychological theory
Donald Olding Hebb · 2005
Cited alongside, same era.
Model compression
Cristian Buciluǎ, Rich Caruana, and Alexandru Niculescu-Mizil · 2006
Cited alongside, same era.
The psychology of the child
Jean Piaget and Barbel Inhelder · 2008
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Cited alongside, same era.
Deep convolutional inverse graphics network
Tejas D Kulkarni, William F Whitney, Pushmeet Kohli, and Josh Tenenbaum · 2015
Later among the works it cites.
Unifying distillation and privileged information
David Lopez-Paz, Léon Bottou, Bernhard Schölkopf, and Vladimir Vapnik · 2015
Later among the works it cites.
Learning using privileged information: Similarity control and knowledge transfer
Vladimir Vapnik and Rauf Izmailov · 2015
Later among the works it cites.
Krzysztof J Geras, Abdel-rahman Mohamed, Rich Caruana, Gregor Urban, Shengjie Wang, Ozlem Aslan, Matthai Philipose, Matthew Richardson, and Charles Sutton · 2015
Later among the works it cites.
Learning multiple tasks with deep relationship networks
Mingsheng Long and Jianmin Wang · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Efficient backprop
Yann A LeCun, Léon Bottou, Genevieve B Orr, and Klaus-Robert Müller · 2012
Cited alongside, same era.
An empirical investigation of catastrophic forgetting in gradient-based neural networks
Ian J Goodfellow, Mehdi Mirza, Da Xiao, Aaron Courville, and Yoshua Bengio · 2013
Cited alongside, same era.
Representation learning: A review and new perspectives
Yoshua Bengio, Aaron Courville, and Pierre Vincent · 2013
Cited alongside, same era.
Hierarchical modular optimization of convolutional networks achieves representations similar to macaque it and human ventral stream
Daniel L Yamins, Ha Hong, Charles Cadieu, and James J DiCarlo · 2013
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Andrew M Saxe, James L McClelland, and Surya Ganguli · 2013
Cited alongside, same era.
How transferable are features in deep neural networks?
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson · 2014
Cited alongside, same era.
Later among the works it cites.
Policy Distillation
A. A. Rusu, S. Gomez Colmenarejo, C. Gulcehre, G. Desjardins, J. Kirkpatrick, R. Pascanu, V. Mnih, K. Kavukcuoglu, and R. Hadsell · 2015
Later among the works it cites.
Actor-mimic: Deep multitask and transfer reinforcement learning
Emilio Parisotto, Jimmy Lei Ba, and Ruslan Salakhutdinov · 2015
Later among the works it cites.
Matconvnet: Convolutional neural networks for matlab
Andrea Vedaldi and Karel Lenc · 2015
Later among the works it cites.
ilab-20m: A large-scale controlled object dataset to investigate deep learning
Borji Ali, Wei Liu, Izadi Saeed, and Itti Laurent · 2016
Closest in time.
Bridging the gaps between residual learning, recurrent neural networks and visual cortex
Qianli Liao and Tomaso Poggio · 2016
Closest in time.
Building machines that learn and think like people
Brenden M Lake, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman · 2016
Closest in time.
Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation
Tejas D Kulkarni, Karthik R Narasimhan, Ardavan Saeedi, and Joshua B Tenenbaum · 2016
Closest in time.
A deep hierarchical approach to lifelong learning in minecraft
Chen Tessler, Shahar Givony, Tom Zahavy, Daniel J Mankowitz, and Shie Mannor · 2016
Closest in time.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Closest in time.
Graying the black box: Understanding dqns
Tom Zahavy, Nir Ben Zrihem, and Shie Mannor · 2016
Closest in time.
Do deep convolutional nets really need to be deep (or even convolutional)?
Gregor Urban, Krzysztof J Geras, Samira Ebrahimi Kahou, Ozlem Aslan, Shengjie Wang, Rich Caruana, Abdelrahman Mohamed, Matthai Philipose, and Matt Richardson · 2016
Closest in time.
Beyond sharing weights for deep domain adaptation
Artem Rozantsev, Mathieu Salzmann, and Pascal Fua · 2016
Closest in time.