Fetching the paper…
Reading the bibliography…
Online Continual Learning (OCL) studies learning over a continuous data stream without observing any single example more than once, a setting that is closer to the experience of humans and systems that must learn "on-the-wild".
Present position and potential developments: Some personal views statistical theory the prequential approach
A Philip Dawid · 1984
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
Michael McCloskey and Neal J Cohen · 1989
Earlier work this paper cites.
Connectionist models of recognition memory: constraints imposed by learning and forgetting functions
Roger Ratcliff · 1990
Earlier work this paper cites.
Adaptive mixtures of local experts
Robert A Jacobs, Michael I Jordan, Steven J Nowlan, and Geoffrey E Hinton · 1991
Earlier work this paper cites.
Complex adaptive systems
John H Holland · 1992
Earlier work this paper cites.
Multitask learning
Rich Caruana · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Products of experts
GE Hinton · 1999
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin · 2003
Earlier work this paper cites.
Europarl: A parallel corpus for statistical machine translation
Philipp Koehn · 2005
Earlier work this paper cites.
The british national corpus, version 3 (bnc xml edition)
BNC Consortium et al · 2007
Earlier work this paper cites.
Findings of the 2009 workshop on statistical machine translation
Chris Callison-Burch, Philipp Koehn, Christof Monz, and Josh Schroeder · 2009
Earlier work this paper cites.
Recurrent neural network based language model
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur · 2010
Earlier work this paper cites.
Learning factored representations in a deep mixture of experts
David Eigen, Marc’Aurelio Ranzato, and Ilya Sutskever · 2013
Earlier work this paper cites.
icub world: Friendly robots help building good vision data-sets
Sean Fanello, Carlo Ciliberto, Matteo Santoro, Lorenzo Natale, Giorgio Metta, Lorenzo Rosasco, and Francesca Odone · 2013
Earlier work this paper cites.
On evaluating stream learning algorithms
João Gama, Raquel Sebastião, and Pedro Pereira Rodrigues · 2013
Earlier work this paper cites.
Who’s that actor? automatic labelling of actors in tv series starting from imdb images
Rahaf Aljundi, Punarjay Chakravarty, and Tinne Tuytelaars · 2016
Earlier work this paper cites.
Andrei A Rusu, Neil C Rabinowitz, Guillaume Desjardins, Hubert Soyer, James Kirkpatrick, Koray Kavukcuoglu, Razvan Pascanu, and Raia Hadsell · 2016
Earlier work this paper cites.
Expert gate: Lifelong learning with a network of experts
Rahaf Aljundi, Punarjay Chakravarty, and Tinne Tuytelaars · 2017
Earlier work this paper cites.
Pathnet: Evolution channels gradient descent in super neural networks
Chrisantha Fernando, Dylan Banarse, Charles Blundell, Yori Zwols, David Ha, Andrei A Rusu, Alexander Pritzel, and Daan Wierstra · 2017
Cited alongside, same era.
Unbounded cache model for online language modeling with open vocabulary
Edouard Grave, Moustapha M Cisse, and Armand Joulin · 2017
Cited alongside, same era.
Learning to create and reuse words in open-vocabulary neural language modeling
Kazuya Kawakami, Chris Dyer, and Phil Blunsom · 2017
Cited alongside, same era.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al · 2017
Cited alongside, same era.
Overcoming catastrophic forgetting by incremental moment matching
Sang-Woo Lee, Jin-Hwa Kim, Jaehyun Jun, Jung-Woo Ha, and Byoung-Tak Zhang · 2017
Cited alongside, same era.
Training tips for the transformer model
Martin Popel and Ondřej Bojar · 2018
Later among the works it cites.
Online deep learning: Learning deep neural networks on the fly
Doyen Sahoo, Quang Pham, Jing Lu, and Steven C.H. Hoi · 2018
Later among the works it cites.
Overcoming catastrophic forgetting with hard attention to the task
Joan Serra, Didac Suris, Marius Miron, and Alexandros Karatzoglou · 2018
Later among the works it cites.
Toward training recurrent neural networks for lifelong learning
Shagun Sodhani, Sarath Chandar, and Yoshua Bengio · 2018
Later among the works it cites.
Breaking the softmax bottleneck: A high-rank rnn language model
Zhilin Yang, Zihang Dai, Ruslan Salakhutdinov, and William W Cohen · 2018
Later among the works it cites.
Episodic memory in lifelong language learning
Cyprien de Masson d’Autume, Sebastian Ruder, Lingpeng Kong, and Dani Yogatama · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Core50: a new dataset and benchmark for continuous object recognition
Vincenzo Lomonaco and Davide Maltoni · 2017
Cited alongside, same era.
Gradient episodic memory for continual learning
David Lopez-Paz and Marc’Aurelio Ranzato · 2017
Cited alongside, same era.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Cited alongside, same era.
Learning multiple visual domains with residual adapters
Sylvestre-Alvise Rebuffi, Hakan Bilen, and Andrea Vedaldi · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Continual learning through synaptic intelligence
Friedemann Zenke, Ben Poole, and Surya Ganguli · 2017
Cited alongside, same era.
Later among the works it cites.
Continual lifelong learning with neural networks: A review
German I Parisi, Ronald Kemker, Jose L Part, Christopher Kanan, and Stefan Wermter · 2019
Later among the works it cites.
Mixture models for diverse machine translation: Tricks of the trade
Tianxiao Shen, Myle Ott, Michael Auli, and Marc’Aurelio Ranzato · 2019
Later among the works it cites.
Continual learning with adaptive weights (claw)
Tameem Adel, Han Zhao, and Richard E Turner · 2020
Closest in time.
Uncertainty-guided continual learning with bayesian neural networks
Sayna Ebrahimi, Mohamed Elhoseiny, Trevor Darrell, and Marcus Rohrbach · 2020
Closest in time.
A neural dirichlet process mixture model for task-free continual learning
Soochan Lee, Junsoo Ha, Dongsu Zhang, and Gunhee Kim · 2020
Closest in time.
Continual learning for robotics: Definition, framework, learning strategies, opportunities and challenges
Timothée Lesort, Vincenzo Lomonaco, Andrei Stoian, Davide Maltoni, David Filliat, and Natalia Díaz-Rodríguez · 2020
Closest in time.
Compositional language continual learning
Yuanpeng Li, Liang Zhao, Kenneth Church, and Mohamed Elhoseiny · 2020
Closest in time.
Online continual learning on sequences
German I Parisi and Vincenzo Lomonaco · 2020
Closest in time.
Stream-51: Streaming classification and novelty detection from videos
Ryne Roady, Tyler L Hayes, Hitesh Vaidya, and Christopher Kanan · 2020
Closest in time.
Functional regularisation for continual learning with gaussian processes
Michalis K Titsias, Jonathan Schwarz, Alexander G de G Matthews, Razvan Pascanu, and Yee Whye Teh · 2020
Closest in time.
Continual learning with hypernetworks
Johannes von Oswald, Christian Henning, João Sacramento, and Benjamin F Grewe · 2020
Closest in time.
Scalable and order-robust continual learning with additive parameter decomposition
Jaehong Yoon, Saehoon Kim, Eunho Yang, and Sung Ju Hwang · 2020
Closest in time.