Fetching the paper…
Reading the bibliography…
Catastrophic forgetting is the notorious vulnerability of neural networks to the change of the data distribution while learning.
Catastrophic interference in connectionist networks: The sequential learning problem
McCloskey, M. and Cohen, N. J · 1989
Earlier work this paper cites.
A practical Bayesian framework for backpropagation networks
MacKay, D. J. C · 1992
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
Hinton, G. E. and Camp, D. V · 1993
Earlier work this paper cites.
Bayesian Learning for Neural Networks
Neal, R. M · 1994
Earlier work this paper cites.
Natural gradient works efficiently in learning
Amari, S.-I · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A. and Hinton, G · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Glorot, X. and Bengio, Y · 2010
Earlier work this paper cites.
Practical variational inference for neural networks
Graves, A · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Netzer, Y., Wang, T., Coates, A., Bissacco, A., Wu, B., and Ng, A. Y · 2011
Earlier work this paper cites.
Bayesian learning via stochastic gradient langevin dynamics
Welling, M. and Teh, Y. W · 2011
Earlier work this paper cites.
Streaming variational bayes
Broderick, T., Boyd, N., Wibisono, A., Wilson, A. C., and Jordan, M. I · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2013
Earlier work this paper cites.
Identifying and attacking the saddle point problem in high-dimensional non-convex optimization
Dauphin, Y., Pascanu, R., Gulcehre, C., Cho, K., Ganguli, S., and Bengio, Y · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A · 2014
Cited alongside, same era.
Expectation backpropagation: Parameter-free training of multilayer neural networks with continuous or discrete weights
Soudry, D., Hubara, I., and Meir, R · 2014
Cited alongside, same era.
Bayesian dark knowledge
Balan, A. K., Rathod, V., Murphy, K. P., and Welling, M · 2015
Cited alongside, same era.
Weight uncertainty in neural network
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D · 2015
Cited alongside, same era.
Probabilistic backpropagation for scalable learning of bayesian neural networks
Hernández-Lobato, J. M. and Adams, R · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Continual learning with deep generative replay
Shin, H., Lee, J. K., Kim, J., and Kim, J · 2017
Later among the works it cites.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
Xiao, H., Rasul, K., and Vollgraf, R · 2017
Later among the works it cites.
Continual learning through synaptic intelligence
Zenke, F., Poole, B., and Ganguli, S · 2017
Later among the works it cites.
Noisy natural gradient as variational inference
Zhang, G., Sun, S., Duvenaud, D., and Grosse, R · 2017
Later among the works it cites.
Memory aware synapses: Learning what (not) to forget
Aljundi, R., Babiloni, F., Elhoseiny, M., Rohrbach, M., and Tuytelaars, T · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ioffe, S. and Szegedy, C · 2015
Cited alongside, same era.
Optimizing neural networks with kronecker-factored approximate curvature
Martens, J. and Grosse, R · 2015
Cited alongside, same era.
Black-box α \alpha -divergence minimization
Hernández-Lobato, J. M., Li, Y., Rowland, M., Hernández-Lobato, D., Bui, T., and Turner, R. E · 2016
Cited alongside, same era.
Rusu, A. A., Rabinowitz, N. C., Desjardins, G., Soyer, H., Kirkpatrick, J., Kavukcuoglu, K., Pascanu, R., and Hadsell, R · 2016
Cited alongside, same era.
Overcoming catastrophic forgetting in neural networks
Kirkpatrick, J., Pascanu, R., Rabinowitz, N., Veness, J., Desjardins, G., Rusu, A. A., Milan, K., Quan, J., Ramalho, T., Grabska-Barwinska, A., et al · 2017
Cited alongside, same era.
Overcoming catastrophic forgetting by incremental moment matching
Lee, S., Kim, J., Ha, J., and Zhang, B · 2017
Cited alongside, same era.
Learning without forgetting
Li, Z. and Hoiem, D · 2017
Cited alongside, same era.
Chaudhry, A., Dokania, P. K., Ajanthan, T., and Torr, P. H · 2018
Closest in time.
Towards robust evaluations of continual learning
Farquhar, S. and Gal, Y · 2018
Closest in time.
Re-evaluating continual learning scenarios: A categorization and case for strong baselines
Hsu, Y.-C., Liu, Y.-C., and Kira, Z · 2018
Closest in time.
Fast and scalable bayesian deep learning by weight-perturbation in adam
Khan, M. E., Nielsen, D., Tangkaratt, V., Lin, W., Gal, Y., and Srivastava, A · 2018
Closest in time.
Continual lifelong learning with neural networks: A review
Parisi, G. I., Kemker, R., Part, J. L., Kanan, C., and Wermter, S · 2018
Closest in time.
Online structured laplace approximations for overcoming catastrophic forgetting
Ritter, H., Botev, A., and Barber, D · 2018
Closest in time.
Overpruning in variational bayesian neural networks
Trippe, B. and Turner, R · 2018
Closest in time.
Generative replay with feedback connections as a general strategy for continual learning
van de Ven, G. M. and Tolias, A. S · 2018
Closest in time.