Fetching the paper…
Reading the bibliography…
In this paper, we propose a multi-timescale replay (MTR) buffer for improving continual learning in RL agents faced with environments that are changing continuously over time at timescales that are unknown to the agent.
Random sampling with a reservoir
Jeffrey S Vitter · 1985
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
M. McCloskey and J. N. Cohen · 1989
Earlier work this paper cites.
On the form of forgetting
John T Wixted and Ebbe B Ebbesen · 1991
Earlier work this paper cites.
Self-improving reactive agents based on reinforcement learning, planning and teaching
Long-Ji Lin · 1992
Earlier work this paper cites.
Continual learning in reinforcement environments
Mark Bishop Ring · 1994
Earlier work this paper cites.
Catastrophic forgetting, rehearsal and pseudorehearsal
Anthony Robins · 1995
Earlier work this paper cites.
One hundred years of forgetting: A quantitative description of retention
David C Rubin and Amy E Wenzel · 1996
Earlier work this paper cites.
Reinforcement learning: An introduction , volume 1
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Estimation of dependences based on empirical data
Vladimir Vapnik · 2006
Earlier work this paper cites.
Ella: An efficient lifelong learning algorithm
P. Ruvolo and E. Eaton · 2013
Earlier work this paper cites.
The importance of experience replay database composition in deep reinforcement learning
T. de Bruin, J. Kober, K. Tuyls, and R. Babuška · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Cited alongside, same era.
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2015
Cited alongside, same era.
https://sites.google.com/site/cldlnips2016/home , 2016
Continual Learning and Deep Networks Workshop · 2016
Cited alongside, same era.
Off-policy experience retention for deep actor-critic learning
T. de Bruin, J. Kober, K. Tuyls, and R. Babuška · 2016
Cited alongside, same era.
https://sites.google.com/view/continual2018/home , 2018
Continual Learning Workshop · 2018
Later among the works it cites.
Stable baselines
Ashley Hill, Antonin Raffin, Maximilian Ernestus, Adam Gleave, Anssi Kanervisto, Rene Traore, Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, and Yuhuai Wu · 2018
Later among the works it cites.
Selective experience replay for lifelong learning
David Isele and Akansel Cosgun · 2018
Later among the works it cites.
Continual reinforcement learning with complex synapses
Christos Kaplanis, Murray Shanahan, and Claudia Clopath · 2018
Later among the works it cites.
Assessing generalization in deep reinforcement learning, 2018
Charles Packer, Katelyn Gao, Jernej Kos, Philipp Krähenbühl, Vladlen Koltun, and Dawn Song · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Note on the power law of forgetting
Michael J Kahana and Mark Adler · 2017
Cited alongside, same era.
Overcoming catastrophic forgetting in neural networks
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska, et al · 2017
Cited alongside, same era.
Gradient episodic memory for continual learning
David Lopez-Paz et al · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Continual learning through synaptic intelligence
F. Zenke, B. Poole, and S. Ganguli · 2017
Cited alongside, same era.
A deeper look at experience replay
S. Zhang and R.S. Sutton · 2017
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine
Cited in the paper.
Martin Arjovsky, Léon Bottou, Ishaan Gulrajani, and David Lopez-Paz · 2019
Later among the works it cites.
Continual learning with tiny episodic memories
Arslan Chaudhry, Marcus Rohrbach, Mohamed Elhoseiny, Thalaiyasingam Ajanthan, Puneet K Dokania, Philip HS Torr, and Marc’Aurelio Ranzato · 2019
Later among the works it cites.
Policy consolidation for continual reinforcement learning
Christos Kaplanis, Murray Shanahan, and Claudia Clopath · 2019
Later among the works it cites.
Experience replay for continual learning
David Rolnick, Arun Ahuja, Jonathan Schwarz, Timothy Lillicrap, and Gregory Wayne · 2019
Later among the works it cites.
Boosting soft actor-critic: Emphasizing recent experience without forgetting the past
Che Wang and Keith Ross · 2019
Later among the works it cites.