Fetching the paper…
Reading the bibliography…
We propose a method for tackling catastrophic forgetting in deep reinforcement learning that is \textit{agnostic} to the timescale of changes in the distribution of experiences, does not require knowledge of task boundaries, and can adapt in \textit{continuously} changing environments.
Catastrophic interference in connectionist networks: The sequential learning problem
McCloskey, M. and Cohen, J. N · 1989
Earlier work this paper cites.
Child: A first step towards continual learning
Ring, M. B · 1997
Earlier work this paper cites.
Learning and memory: An integrated approach
Anderson, J. R · 2000
Earlier work this paper cites.
Principles derived from the study of simple skills do not generalize to complex skill learning
Wulf, G. and Shea, C. H · 2002
Earlier work this paper cites.
Ella: An efficient lifelong learning algorithm
Ruvolo, P. and Eaton, E · 2013
Earlier work this paper cites.
Distilling the knowledge in a neural network
Hinton, G., Vinyals, O., and Dean, J · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Lillicrap, T. P., Hunt, J. J., Pritzel, A., Heess, N., Erez, T., Tassa, Y., Silver, D., and Wierstra, D · 2015
Earlier work this paper cites.
Rusu, A. A., Colmenarejo, S. G., Gulcehre, C., Desjardins, G., Kirkpatrick, J., Pascanu, R., Mnih, V., Kavukcuoglu, K., and Hadsell, R · 2015
Earlier work this paper cites.
High-dimensional continuous control using generalized advantage estimation
Schulman, J., Moritz, P., Levine, S., Jordan, M., and Abbeel, P · 2015
Earlier work this paper cites.
Computational principles of synaptic memory consolidation
Benna, M. K. and Fusi, S · 2016
Earlier work this paper cites.
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W · 2016
Earlier work this paper cites.
Active long term memory networks
Furlanello, T., Zhao, J., Saxe, A. M., Itti, L., and Tjan, B. S · 2016
Earlier work this paper cites.
Deep reinforcement learning from self-play in imperfect-information games
Heinrich, J. and Silver, D · 2016
Cited alongside, same era.
The forget-me-not process
Milan, K., Veness, J., Kirkpatrick, J., Bowling, M., Koop, A., and Hassabis, D · 2016
Cited alongside, same era.
Rusu, A. A., Rabinowitz, N. C., Desjardins, G., Soyer, H., Kirkpatrick, J., Kavukcuoglu, K., Pascanu, R., and Hadsell, R · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al · 2016
Cited alongside, same era.
Continuous adaptation via meta-learning in nonstationary and competitive environments
Al-Shedivat, M., Bansal, T., Burda, Y., Sutskever, I., Mordatch, I., and Abbeel, P · 2017
Distral: Robust multitask reinforcement learning
Teh, Y., Bapst, V., Czarnecki, W. M., Quan, J., Kirkpatrick, J., Hadsell, R., Heess, N., and Pascanu, R · 2017
Later among the works it cites.
Continual learning through synaptic intelligence
Zenke, F., Poole, B., and Ganguli, S · 2017
Later among the works it cites.
Online incremental feature learning with denoising autoencoders
Zhou, G., Sohn, K., and Lee, H · 2017
Later among the works it cites.
Continuous adaptation via meta-learning in nonstationary and competitive environments
Al-Shedivat, M., Bansal, T., Burda, Y., Sutskever, I., Mordatch, I., and Abbeel, P · 2018
Later among the works it cites.
Emergent complexity via multi-agent competition
Bansal, T., Pachocki, J., Sidor, S., Sutskever, I., and Mordatch, I · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Openai baselines
Dhariwal, P., Hesse, C., Klimov, O., Nichol, A., Plappert, M., Radford, A., Schulman, J., Sidor, S., Wu, Y., and Zhokhov, P · 2017
Cited alongside, same era.
Benchmark environments for multitask learning in continuous domains
Henderson, P., Chang, W.-D., Shkurti, F., Hansen, J., Meger, D., and Dudek, G · 2017
Cited alongside, same era.
Overcoming catastrophic forgetting in neural networks
Kirkpatrick, J., Pascanu, R., Rabinowitz, N., Veness, J., Desjardins, G., Rusu, A. A., Milan, K., Quan, J., Ramalho, T., Grabska-Barwinska, A., et al · 2017
Cited alongside, same era.
Learning without forgetting
Li, Z. and Hoiem, D · 2017
Cited alongside, same era.
Gradient episodic memory for continual learning
Lopez-Paz, D. et al · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Cited alongside, same era.
Continual learning with deep generative replay
Shin, H., Lee, J. K., Kim, J., and Kim, J · 2017
Cited alongside, same era.
Isele, D. and Cosgun, A · 2018
Later among the works it cites.
Continual reinforcement learning with complex synapses
Kaplanis, C., Shanahan, M., and Clopath, C · 2018
Later among the works it cites.
Experience replay for continual learning
Rolnick, D., Ahuja, A., Schwarz, J., Lillicrap, T. P., and Wayne, G · 2018
Later among the works it cites.
Progress & compress: A scalable framework for continual learning
Schwarz, J., Luketina, J., Czarnecki, W. M., Grabska-Barwinska, A., Teh, Y. W., Pascanu, R., and Hadsell, R · 2018
Later among the works it cites.
Memory-based parameter adaptation
Sprechmann, P., Jayakumar, S. M., Rae, J. W., Pritzel, A., Badia, A. P., Uria, B., Vinyals, O., Hassabis, D., Pascanu, R., and Blundell, C · 2018
Later among the works it cites.
Xu, J. and Zhu, Z · 2018
Later among the works it cites.