Fetching the paper…
Reading the bibliography…
Learning a sequence of tasks without access to i.i.d.
Self-improving reactive agents based on reinforcement learning, planning and teaching
L.-J. Lin · 1992
Earlier work this paper cites.
A practical bayesian framework for backpropagation networks
D. J. MacKay · 1992
Earlier work this paper cites.
Learning to Control Fast-Weight Memories: An Alternative to Dynamic Recurrent Networks
J. Schmidhuber · 1992
Earlier work this paper cites.
Flat minima
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Information theory, inference and learning algorithms
D. J. MacKay · 2003
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
A kernel two-sample test
A. Gretton, K. M. Borgwardt, M. J. Rasch, B. Schölkopf, and A. Smola · 2012
Earlier work this paper cites.
An empirical investigation of catastrophic forgetting in gradient-based neural networks
I. J. Goodfellow, M. Mirza, D. Xiao, A. Courville, and Y. Bengio · 2013
Earlier work this paper cites.
Cramér-rao lower bound and information geometry
F. Nielsen · 2013
Earlier work this paper cites.
Revisiting natural gradient for deep networks
R. Pascanu and Y. Bengio · 2013
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2014
Earlier work this paper cites.
Coresets for nonparametric estimation - the case of dp-means
O. Bachem, M. Lucic, and A. Krause · 2015
Earlier work this paper cites.
Weight uncertainty in neural networks
C. Blundell, J. Cornebise, K. Kavukcuoglu, and D. Wierstra · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network
G. Hinton, O. Vinyals, and J. Dean · 2015
Earlier work this paper cites.
Variational dropout and the local reparameterization trick
D. P. Kingma, T. Salimans, and M. Welling · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks
D. Hendrycks and K. Gimpel · 2016
Earlier work this paper cites.
A kernelized stein discrepancy for goodness-of-fit tests
Q. Liu, J. Lee, and M. Jordan · 2016
Earlier work this paper cites.
f-gan: Training generative neural samplers using variational divergence minimization
S. Nowozin, B. Cseke, and R. Tomioka · 2016
Earlier work this paper cites.
Wide residual networks
S. Zagoruyko and N. Komodakis · 2016
Earlier work this paper cites.
Towards principled methods for training generative adversarial networks
M. Arjovsky and L. Bottou · 2017
Earlier work this paper cites.
Variational inference: A review for statisticians
D. M. Blei, A. Kucukelbir, and J. D. McAuliffe · 2017
Earlier work this paper cites.
Hypernetworks
D. Ha, A. Dai, and Q. Le · 2017
Earlier work this paper cites.
Variational inference using implicit distributions
F. Huszár · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska, D. Hassabis, C. Clopath, D. Kumaran, and R. Hadsell · 2017
Earlier work this paper cites.
Bayesian Hypernetworks
D. Krueger, C.-W. Huang, R. Islam, R. Turner, A. Lacoste, and A. Courville · 2017
Earlier work this paper cites.
Simple and scalable predictive uncertainty estimation using deep ensembles
B. Lakshminarayanan, A. Pritzel, and C. Blundell · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting by incremental moment matching
S.-W. Lee, J.-H. Kim, J. Jun, J.-W. Ha, and B.-T. Zhang · 2017
Earlier work this paper cites.
Multiplicative Normalizing Flows for Variational Bayesian Neural Networks
C. Louizos and M. Welling · 2017
Earlier work this paper cites.
Adversarial variational bayes: Unifying variational autoencoders and generative adversarial networks
L. Mescheder, S. Nowozin, and A. Geiger · 2017
Earlier work this paper cites.
Implicit weight uncertainty in neural networks
N. Pawlowski, A. Brock, M. C. Lee, M. Rajchl, and B. Glocker · 2017
Cited alongside, same era.
icarl: Incremental classifier and representation learning
S.-A. Rebuffi, A. Kolesnikov, G. Sperl, and C. H. Lampert · 2017
Cited alongside, same era.
Continual Learning with Deep Generative Replay
H. Shin, J. K. Lee, J. Kim, and J. Kim · 2017
Cited alongside, same era.
Continual Learning Through Synaptic Intelligence
F. Zenke, B. Poole, and S. Ganguli · 2017
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
C. Zhang, S. Bengio, M. Hardt, B. Recht, and O. Vinyals · 2017
Cited alongside, same era.
Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning
S. Depeweg, J.-M. Hernandez-Lobato, F. Doshi-Velez, and S. Udluft · 2018
Cited alongside, same era.
Discriminative representation loss (drl): Connecting deep metric learning to continual learning
Y. Chen, T. Diethe, and P. Flach · 2020
Later among the works it cites.
Continual prototype evolution: Learning online from non-stationary data streams
M. De Lange and T. Tuytelaars · 2020
Later among the works it cites.
Hypermodels for exploration
V. Dwaracherla, X. Lu, M. Ibrahimi, I. Osband, Z. Wen, and B. V. Roy · 2020
Later among the works it cites.
Radial bayesian neural networks: Beyond discrete support in large-scale bayesian deep learning
S. Farquhar, M. A. Osborne, and Y. Gal · 2020
Later among the works it cites.
Meta-consolidation for continual learning
J. K J and V. N Balasubramanian · 2020
Later among the works it cites.
Hierarchical gaussian process priors for bayesian neural network weights
T. Karaletsos and T. D. Bui · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Approximating the predictive distribution via adversarially-trained hypernetworks
C. Henning, J. von Oswald, J. Sacramento, S. C. Surace, J.-P. Pfister, and B. F. Grewe · 2018
Cited alongside, same era.
Stein neural sampler
T. Hu, Z. Chen, H. Sun, J. Bai, M. Ye, and G. Cheng · 2018
Cited alongside, same era.
Note on the quadratic penalties in elastic weight consolidation
F. Huszár · 2018
Cited alongside, same era.
Probabilistic meta-representations of neural networks
T. Karaletsos, P. Dayan, and Z. Ghahramani · 2018
Cited alongside, same era.
Training confidence-calibrated classifiers for detecting out-of-distribution samples
K. Lee, H. Lee, K. Lee, and J. Shin · 2018
Cited alongside, same era.
Gradient estimators for implicit models
Y. Li and R. E. Turner · 2018
Cited alongside, same era.
Later among the works it cites.
Continual learning with bayesian neural networks for non-stationary data
R. Kurle, B. Cseke, A. Klushyn, P. van der Smagt, and S. Günnemann · 2020
Later among the works it cites.
Simple and principled uncertainty estimation with deterministic deep learning via distance awareness
B. Lakshminarayanan, D. Tran, J. Liu, S. Padhy, T. Bedrax-Weiss, and Z. Lin · 2020
Later among the works it cites.
A neural dirichlet process mixture model for task-free continual learning
S. Lee, J. Ha, D. Zhang, and G. Kim · 2020
Later among the works it cites.
Ensemble distribution distillation
A. Malinin, B. Mlodozeniec, and M. Gales · 2020
Later among the works it cites.
New insights and perspectives on the natural gradient method
J. Martens · 2020
Later among the works it cites.
Class-incremental learning: survey and performance evaluation on image classification
M. Masana, X. Liu, B. Twardowski, M. Menta, A. D. Bagdanov, and J. van de Weijer · 2020
Later among the works it cites.
A wholistic view of continual learning with deep neural networks: Forgotten lessons and the bridge to active and open world learning
M. Mundt, Y. W. Hong, I. Pliushch, and V. Ramesh · 2020
Later among the works it cites.
Continual deep learning by functional regularisation of memorable past
P. Pan, S. Swaroop, A. Immer, R. Eschenhagen, R. Turner, and M. E. E. Khan · 2020
Later among the works it cites.
Gdumb: A simple approach that questions our progress in continual learning
A. Prabhu, P. H. Torr, and P. K. Dokania · 2020
Later among the works it cites.
Functional regularisation for continual learning with gaussian processes
M. K. Titsias, J. Schwarz, A. G. de G. Matthews, R. Pascanu, and Y. W. Teh · 2020
Later among the works it cites.
Brain-inspired replay for continual learning with artificial neural networks
G. M. van de Ven, H. T. Siegelmann, and A. S. Tolias · 2020
Later among the works it cites.
Continual learning with hypernetworks
J. von Oswald, C. Henning, J. Sacramento, and B. F. Grewe · 2020
Later among the works it cites.
How good is the Bayes posterior in deep neural networks really?
F. Wenzel, K. Roth, B. Veeling, J. Swiatkowski, L. Tran, S. Mandt, J. Snoek, T. Salimans, R. Jenatton, and S. Nowozin · 2020
Later among the works it cites.
Bayesian deep learning and a probabilistic perspective of generalization
A. G. Wilson and P. Izmailov · 2020
Later among the works it cites.
Supermasks in superposition
M. Wortsman, V. Ramanujan, R. Liu, A. Kembhavi, M. Rastegari, J. Yosinski, and A. Farhadi · 2020
Later among the works it cites.
Uncertainty-based out-of-distribution detection requires suitable function space priors
F. D’Angelo and C. Henning · 2021
Closest in time.
Continual learning in recurrent neural networks
B. Ehret, C. Henning, M. R. Cervera, A. Meulemans, J. von Oswald, and B. F. Grewe · 2021
Closest in time.
Variational beam search for novelty detection
A. Li, A. J. Boyd, P. Smyth, and S. Mandt · 2021
Closest in time.
Generalized variational continual learning
N. Loo, S. Swaroop, and R. E. Turner · 2021
Closest in time.
Deterministic neural networks with appropriate inductive biases capture epistemic and aleatoric uncertainty
J. Mukhoti, A. Kirsch, J. van Amersfoort, P. H. Torr, and Y. Gal · 2021
Closest in time.
Neural networks with late-phase weights
J. V. Oswald, S. Kobayashi, J. Sacramento, A. Meulemans, C. Henning, and B. F. Grewe · 2021
Closest in time.
Continual learning without knowing task identities: Do simple models work?
T. Tuor, S. Wang, and K. Leung · 2021
Closest in time.
Class-incremental learning with generative classifiers
G. M. van de Ven, Z. Li, and A. S. Tolias · 2021
Closest in time.
Efficient feature transformations for discriminative and generative continual learning
V. K. Verma, K. J. Liang, N. Mehta, P. Rai, and L. Carin · 2021
Closest in time.
Task-agnostic continual learning using online variational bayes with fixed-point updates
C. Zeno, I. Golan, E. Hoffer, and D. Soudry · 2021
Closest in time.