Fetching the paper…
Reading the bibliography…
One major obstacle towards AI is the poor ability of models to solve new problems quicker, and without forgetting previously acquired knowledge.
Duality in quadratic programming
W. S. Dorn · 1960
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
M. McCloskey and N. J. Cohen · 1989
Earlier work this paper cites.
Connectionist models of recognition memory: Constraints imposed by learning and forgetting functions
R. Ratcliff · 1990
Earlier work this paper cites.
Continual Learning in Reinforcement Environments
M. B. Ring · 1994
Earlier work this paper cites.
A lifelong learning perspective for mobile robot control
S. Thrun · 1994
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: insights from the successes and failures of connectionist models of learning and memory
J. L. McClelland, B. L. McNaughton, and R. C. O’reilly · 1995
Earlier work this paper cites.
Is learning the n-th thing any easier than learning the first?
S. Thrun · 1996
Earlier work this paper cites.
CHILD: A first step towards continual learning
M. B. Ring · 1997
Earlier work this paper cites.
Multitask learning
R. Caruana · 1998
Earlier work this paper cites.
The MNIST database of handwritten digits, 1998
Y. LeCun, C. Cortes, and C. J. Burges · 1998
Earlier work this paper cites.
Lifelong learning algorithms
S. Thrun · 1998
Earlier work this paper cites.
Statistical learning theory
V. Vapnik · 1998
Earlier work this paper cites.
Catastrophic forgetting in connectionist networks
R. M. French · 1999
Earlier work this paper cites.
A model of inductive bias learning
J. Baxter · 2000
Earlier work this paper cites.
Learning with kernels: support vector machines, regularization, optimization, and beyond
B. Scholkopf and A. J. Smola · 2001
Earlier work this paper cites.
A Bayesian approach to unsupervised one-shot learning of object categories
L. Fei-Fei, R. Fergus, and P. Perona · 2003
Earlier work this paper cites.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky · 2009
Earlier work this paper cites.
Learning to detect unseen object classes by between-class attribute transfer
C. Lampert, H. Nickisch, and S. Harmeling · 2009
Cited alongside, same era.
Zero-shot learning with semantic output codes
M. Palatucci, D. A. Pomerleau, G. E. Hinton, and T. Mitchell · 2009
Cited alongside, same era.
A theory of learning from different domains
S. Ben-David, J. Blitzer, K. Crammer, A. Kulesza, F. Pereira, and J. Wortman Vaughan · 2010
Cited alongside, same era.
Toward an architecture for never-ending language learning
A. Carlson, J. Betteridge, B. Kisiel, B. Settles, E. R. Hruschka, and T. M. Mitchell · 2010
Cited alongside, same era.
A survey on transfer learning
S. J. Pan and Q. Yang · 2010
Cited alongside, same era.
Horde: A scalable real-time architecture for learning knowledge from unsupervised sensorimotor interaction
R. S. Sutton, J. Modayil, M. Delp, T. Degris, P. M. Pilarski, A. White, and D. Precup · 2011
Expert gate: Lifelong learning with a network of experts
R. Aljundi, P. Chakravarty, and T. Tuytelaars · 2016
Later among the works it cites.
Learning feed-forward one-shot learners
L. Bertinetto, J. Henriques, J. Valmadre, P. Torr, and A. Vedaldi · 2016
Later among the works it cites.
Less-forgetting Learning in Deep Neural Networks
H. Jung, J. Ju, M. Jung, and J. Kim · 2016
Later among the works it cites.
Learning without forgetting
Z. Li and D. Hoiem · 2016
Later among the works it cites.
Lifelong learning with weighted majority votes
A. Pentina and R. Urner · 2016
Later among the works it cites.
Causal inference by using invariant prediction: identification and confidence intervals
J. Peters, P. Bühlmann, and N. Meinshausen · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning to learn
S. Thrun and L. Pratt · 2012
Cited alongside, same era.
An Empirical Investigation of Catastrophic Forgetting in Gradient-Based Neural Networks
I. J. Goodfellow, M. Mirza, D. Xiao, A. Courville, and Y. Bengio · 2013
Cited alongside, same era.
ELLA: An Efficient Lifelong Learning Algorithm
P. Ruvolo and E. Eaton · 2013
Cited alongside, same era.
Learning factored representations in a deep mixture of experts
D. Eigen, I. Sutskever, and M. Ranzato · 2014
Cited alongside, same era.
Learning and transferring mid-level image representations using convolutional neural networks
M. Oquab, L. Bottou, I. Laptev, and J. Sivic · 2014
Cited alongside, same era.
Efficient representations for lifelong learning and autoencoding
M.-F. Balcan, A. Blum, and S. Vempola · 2015
Cited alongside, same era.
Progressive neural networks
A. A. Rusu, N. C. Rabinowitz, G. Desjardins, H. Soyer, J. Kirkpatrick, K. Kavukcuoglu, R. Pascanu, and R. Hadsell · 2016
Later among the works it cites.
One-shot learning with memory-augmented neural networks
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. Lillicrap · 2016
Later among the works it cites.
Causal and statistical learning
B. Schölkopf, D. Janzing, and D. Lopez-Paz · 2016
Later among the works it cites.
Matching networks for one shot learning
O. Vinyals, C. Blundell, T. Lillicrap, and D. Wierstra · 2016
Later among the works it cites.
CommAI: Evaluating the first steps towards a useful general AI
M. Baroni, A. Joulin, A. Jabri, G. Kruszewski, A. Lazaridou, K. Simonic, and T. Mikolov · 2017
Closest in time.
PathNet: Evolution channels gradient descent in super neural networks
C. Fernando, D. Banarse, C. Blundell, Y. Zwols, D. Ha, A. A. Rusu, A. Pritzel, and D. Wierstra · 2017
Closest in time.
Overcoming catastrophic forgetting in neural networks
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska, et al · 2017
Closest in time.
Training Mixture Models at Scale via Coresets
M. Lucic, M. Faulkner , A. Krause, and D. Feldman · 2017
Closest in time.
Encoder Based Lifelong Learning
A. Rannen Triki, R. Aljundi, M. B. Blaschko, and T. Tuytelaars · 2017
Closest in time.
iCaRL: Incremental classifier and representation learning
S.-A. Rebuffi, A. Kolesnikov, G. Sperl, and C. H. Lampert · 2017
Closest in time.
Improved multitask learning through synaptic intelligence
F. Zenke, B. Poole, and S. Ganguli · 2017
Closest in time.