Fetching the paper…
Reading the bibliography…
This paper develops variational continual learning (VCL), a simple but general framework for continual learning that fuses online variational inference (VI) and recent advances in Monte Carlo VI for neural networks.
Stochastic models, estimation, and control
Peter S. Maybeck · 1982
Earlier work this paper cites.
Clustering to minimize the maximum intercluster distance
Teofilo F. Gonzalez · 1985
Earlier work this paper cites.
A case study of incremental concept induction
Jeffrey C. Schlimmer and Douglas Fisher · 1986
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
Michael McCloskey and Neal J. Cohen · 1989
Earlier work this paper cites.
Training multilayer perceptrons with the extended Kalman algorithm
Sharad Singhal and Lance Wu · 1989
Earlier work this paper cites.
Connectionist models of recognition memory: Constraints imposed by learning and forgetting functions
Roger Ratcliff · 1990
Earlier work this paper cites.
A practical Bayesian framework for backpropagation networks
David J.C. MacKay · 1992
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
Geoffrey E. Hinton and Drew Van Camp · 1993
Earlier work this paper cites.
Online learning with random representations
Richard S. Sutton and Steven D. Whitehead · 1993
Earlier work this paper cites.
CHILD: A first step towards continual learning
Mark B. Ring · 1997
Earlier work this paper cites.
Ensemble learning in Bayesian neural networks
David Barber and Christopher M. Bishop · 1998
Earlier work this paper cites.
Sequential Monte Carlo methods for dynamic systems
Jun S. Liu and Rong Chen · 1998
Earlier work this paper cites.
Sequential Monte Carlo methods to train neural network models
Nando de Freitas, Mahesan Niranjan, Andrew H. Gee, and Arnaud Doucet · 2000
Earlier work this paper cites.
Online variational Bayesian learning
Zoubin Ghahramani and H. Attias · 2000
Earlier work this paper cites.
Online model selection based on the variational Bayes
Masa-Aki Sato · 2001
Earlier work this paper cites.
Task clustering and gating for Bayesian multitask learning
Bart Bakker and Tom Heskes · 2003
Earlier work this paper cites.
Laplace propagation
Alex J. Smola, S.V.N. Vishwanathan, and Eleazar Eskin · 2004
Earlier work this paper cites.
The variational Gaussian approximation revisited
Manfred Opper and Cédric Archambeau · 2009
Cited alongside, same era.
Practical variational inference for neural networks
Alex Graves · 2011
Cited alongside, same era.
Streaming variational Bayes
Tamara Broderick, Nicholas Boyd, Andre Wibisono, Ashia C. Wilson, and Michael I. Jordan · 2013
Cited alongside, same era.
Fixed-form variational posterior approximation through stochastic linear regression
Tim Salimans and David A. Knowles · 2013
Cited alongside, same era.
Stochastic gradient VB and the variational auto-encoder
Diederik P. Kingma and Max Welling · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
Danilo J. Rezende, Shakir Mohamed, and Daan Wierstra · 2014
Cited alongside, same era.
Black-box α \alpha -divergence minimization
José Miguel Hernández-Lobato, Yingzhen Li, Mark Rowland, Daniel Hernández-Lobato, Thang D. Bui, and Richard E. Turner · 2016
Later among the works it cites.
Coresets for scalable Bayesian logistic regression
Jonathan Huggins, Trevor Campbell, and Tamara Broderick · 2016
Later among the works it cites.
Improved variational inference with inverse autoregressive flow
Diederik P. Kingma, Tim Salimans, Rafal Jozefowicz, Xi Chen, Ilya Sutskever, and Max Welling · 2016
Later among the works it cites.
Learning without forgetting
Zhizhong Li and Derek Hoiem · 2016
Later among the works it cites.
Andrei A. Rusu, Neil C. Rabinowitz, Guillaume Desjardins, Hubert Soyer, James Kirkpatrick, Koray Kavukcuoglu, Razvan Pascanu, and Raia Hadsell · 2016
Later among the works it cites.
Generating videos with scene dynamics
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning hidden unit contributions for unsupervised speaker adaptation of neural network acoustic models
Pawel Swietojanski and Steve Renals · 2014
Cited alongside, same era.
Coresets for nonparametric estimation – the case of DP-means
Olivier Bachem, Mario Lucic, and Andreas Krause · 2015
Cited alongside, same era.
Weight uncertainty in neural network
Charles Blundell, Julien Cornebise, Koray Kavukcuoglu, and Daan Wierstra · 2015
Cited alongside, same era.
A recurrent latent variable model for sequential data
Junyoung Chung, Kyle Kastner, Laurent Dinh, Kratarth Goel, Aaron C. Courville, and Yoshua Bengio · 2015
Cited alongside, same era.
Probabilistic backpropagation for scalable learning of Bayesian neural networks
José Miguel Hernández-Lobato and Ryan P. Adams · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2015
Cited alongside, same era.
Carl Vondrick, Hamed Pirsiavash, and Antonio Torralba · 2016
Later among the works it cites.
Streaming sparse Gaussian process approximations
Thang D. Bui, Cuong V. Nguyen, and Richard E. Turner · 2017
Closest in time.
Comment on “Overcoming catastrophic forgetting in NNs”: Are multiple penalties needed?, 2017
Ferenc Huszár · 2017
Closest in time.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A. Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell · 2017
Closest in time.
Overcoming catastrophic forgetting by incremental moment matching
Sang-Woo Lee, Jin-Hwa Kim, Jung-Woo Ha, and Byoung-Tak Zhang · 2017
Closest in time.
Gradient episodic memory for continual learning
David Lopez-Paz and Marc’Aurelio Ranzato · 2017
Closest in time.
Learning multiple visual domains with residual adapters
Sylvestre-Alvise Rebuffi, Hakan Bilen, and Andrea Vedaldi · 2017
Closest in time.
Continual learning in generative adversarial nets
Ari Seff, Alex Beatson, Daniel Suo, and Han Liu · 2017
Closest in time.
Continual learning through synaptic intelligence
Friedemann Zenke, Ben Poole, and Surya Ganguli · 2017
Closest in time.
Note on the quadratic penalties in elastic weight consolidation
Ferenc Huszár · 2018
Closest in time.
Reply to Huszár: The elastic weight consolidation penalty is empirically valid
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A. Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell · 2018
Closest in time.