Fetching the paper…
Reading the bibliography…
Neural networks have achieved remarkable success in many cognitive tasks.
Catastrophic interference in connectionist networks: The sequential learning problem
M. McCloskey and N. J. Cohen · 1989
Earlier work this paper cites.
Theoretical models of learning to learn
J. Baxter · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Dropout as an implicit gating mechanism for continual learning
S.-I. Mirzadeh, M. Farajtabar, and H. Ghasemzadeh · 2004
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
G. E. Hinton, S. Osindero, and Y.-W. Teh · 2006
Earlier work this paper cites.
Understanding the role of training regimes in continual learning
S. I. Mirzadeh, M. Farajtabar, R. Pascanu, and H. Ghasemzadeh · 2006
Earlier work this paper cites.
Cubic regularization of newton method and its global performance
Y. Nesterov and B. T. Polyak · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
Theoretical statistics: Topics for a core course
R. W. Keener · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
An empirical investigation of catastrophic forgetting in gradient-based neural networks
I. J. Goodfellow, M. Mirza, D. Xiao, A. Courville, and Y. Bengio · 2013
Earlier work this paper cites.
ELLA: An efficient lifelong learning algorithm
P. Ruvolo and E. Eaton · 2013
Earlier work this paper cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
A. M. Saxe, J. L. McClelland, and S. Ganguli · 2013
Earlier work this paper cites.
Convex optimization: Algorithms and complexity
S. Bubeck · 2014
Earlier work this paper cites.
Qualitatively characterizing neural network optimization problems
I. J. Goodfellow, O. Vinyals, and A. M. Saxe · 2014
Earlier work this paper cites.
Error-driven incremental learning in deep convolutional neural network for large-scale image classification
T. Xiao, J. Zhang, K. Yang, Y. Peng, and Z. Zhang · 2014
Earlier work this paper cites.
The loss surfaces of multilayer networks
A. Choromanska, M. Henaff, M. Mathieu, G. B. Arous, and Y. LeCun · 2015
Earlier work this paper cites.
Optimizing neural networks with Kronecker-factored approximate curvature
J. Martens and R. Grosse · 2015
Cited alongside, same era.
Multi-task and lifelong learning of kernels
A. Pentina and S. Ben-David · 2015
Cited alongside, same era.
A Kronecker-factored approximate fisher matrix for convolution layers
R. Grosse and J. Martens · 2016
Cited alongside, same era.
Continual learning through evolvable neural turing machines
B. Lüders, M. Schläger, and S. Risi · 2016
Cited alongside, same era.
The benefit of multitask representation learning
A. Maurer, M. Pontil, and B. Romera-Paredes · 2016
Cited alongside, same era.
A. A. Rusu, N. C. Rabinowitz, G. Desjardins, H. Soyer, J. Kirkpatrick, K. Kavukcuoglu, R. Pascanu, and R. Hadsell · 2016
Uniform convergence of gradients for non-convex learning and optimization
D. J. Foster, A. Sekhari, and K. Sridharan · 2018
Later among the works it cites.
Neural tangent kernel: Convergence and generalization in neural networks
A. Jacot, F. Gabriel, and C. Hongler · 2018
Later among the works it cites.
Foundations of machine learning
M. Mohri, A. Rostamizadeh, and A. Talwalkar · 2018
Later among the works it cites.
Closed-loop GAN for continual learning
A. Rios and L. Itti · 2018
Later among the works it cites.
Online structured Laplace approximations for overcoming catastrophic forgetting
H. Ritter, A. Botev, and D. Barber · 2018
Later among the works it cites.
Lifelong learning with dynamically expandable networks
J. Yoon, E. Yang, J. Lee, and S. J. Hwang · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Regret bounds for lifelong learning
P. Alquier, M. Pontil, et al · 2017
Cited alongside, same era.
How to escape saddle points efficiently
C. Jin, R. Ge, P. Netrapalli, S. M. Kakade, and M. I. Jordan · 2017
Cited alongside, same era.
Deep generative dual memory network for continual learning
N. Kamra, U. Gupta, and Y. Liu · 2017
Cited alongside, same era.
Overcoming catastrophic forgetting in neural networks
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska, et al · 2017
Cited alongside, same era.
Gradient episodic memory for continual learning
D. Lopez-Paz and M. Ranzato · 2017
Cited alongside, same era.
Variational continual learning
C. V. Nguyen, Y. Li, T. D. Bui, and R. E. Turner · 2017
Cited alongside, same era.
Later among the works it cites.
A unifying Bayesian view of continual learning
S. Farquhar and Y. Gal · 2019
Later among the works it cites.
C. Finn, A. Rajeswaran, S. Kakade, and S. Levine · 2019
Later among the works it cites.
Reconciling meta-learning and continual learning with online mixtures of tasks
G. Jerfel, E. Grant, T. L. Griffiths, and K. A. Heller · 2019
Later among the works it cites.
Learn to grow: A continual structure learning framework for overcoming catastrophic forgetting
X. Li, Y. Zhou, T. Wu, R. Socher, and C. Xiong · 2019
Later among the works it cites.
Toward understanding catastrophic forgetting in continual learning
C. V. Nguyen, A. Achille, M. Lam, T. Hassner, V. Mahadevan, and S. Soatto · 2019
Later among the works it cites.
Continual learning by asymmetric loss approximation with single-side overestimation
D. Park, S. Hong, B. Han, and K. M. Lee · 2019
Later among the works it cites.
Functional regularisation for continual learning using Gaussian Processes
M. K. Titsias, J. Schwarz, A. G. d. G. Matthews, R. Pascanu, and Y. W. Teh · 2019
Later among the works it cites.
Lifelong optimization with low regret
Y.-S. Wu, P.-A. Wang, and C.-J. Lu · 2019
Later among the works it cites.
Prototype reminding for continual learning
M. Zhang, T. Wang, J. H. Lim, and J. Feng · 2019
Later among the works it cites.
Generalisation guarantees for continual learning with orthogonal gradient descent
M. A. Bennani and M. Sugiyama · 2020
Closest in time.
Orthogonal gradient descent for continual learning
M. Farajtabar, N. Azizan, A. Mott, and A. Li · 2020
Closest in time.