Fetching the paper…
Reading the bibliography…
One approach to meet the challenges of deep lifelong reinforcement learning (LRL) is careful management of the agent's learning experiences, to learn (without forgetting) and build internal meta-models (of the tasks, environments, agents, and world).
Continual learning using world models for pseudo-rehearsal
Nicholas Ketz, Soheil Kolouri, and Praveen Pilly · 1903
Earlier work this paper cites.
Using World Models for Pseudo-Rehearsal in Continual Learning
Nicholas Ketz, Soheil Kolouri, and Praveen Pilly · 1903
Earlier work this paper cites.
Unsupervised progressive learning and the stam architecture
James Smith, Cameron Taylor, Seth Baer, and Constantine Dovrolis · 1904
Earlier work this paper cites.
Three scenarios for continual learning
Gido M. van de Ven and Andreas S. Tolias · 1904
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Sample complexity of multi-task reinforcement learning
Emma Brunskill and Lihong Li · 2013
Earlier work this paper cites.
Playing Atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Earlier work this paper cites.
Auto-Encoding Variational Bayes
Diederik P. Kingma and Max Welling · 2014
Earlier work this paper cites.
Andrei A Rusu, Sergio Gomez Colmenarejo, Caglar Gulcehre, Guillaume Desjardins, James Kirkpatrick, Razvan Pascanu, Volodymyr Mnih, Koray Kavukcuoglu, and Raia Hadsell · 2015
Earlier work this paper cites.
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2015
Earlier work this paper cites.
Hidden technical debt in machine learning systems
David Sculley, Gary Holt, Daniel Golovin, Eugene Davydov, Todd Phillips, Dietmar Ebner, Vinay Chaudhary, Michael Young, Jean-Francois Crespo, and Dan Dennison · 2015
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A. Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell · 2016
Earlier work this paper cites.
Conditional image generation with pixelcnn decoders
Aaron Van den Oord, Nal Kalchbrenner, Lasse Espeholt, Oriol Vinyals, Alex Graves, et al · 2016
Earlier work this paper cites.
icarl: Incremental classifier and representation learning
Sylvestre-Alvise Rebuffi, Alexander Kolesnikov, Georg Sperl, and Christoph H Lampert · 2017
Earlier work this paper cites.
Multi-agent distributed lifelong learning for collective knowledge acquisition
Mohammad Rostami, Soheil Kolouri, Kyungnam Kim, and Eric Eaton · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Continual learning with deep generative replay
Hanul Shin, Jung Kwon Lee, Jaehong Kim, and Jiwon Kim · 2017
Earlier work this paper cites.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Cited alongside, same era.
Starcraft ii: A new challenge for reinforcement learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John Agapiou, Julian Schrittwieser, et al · 2017
Cited alongside, same era.
A deeper look at experience replay
Shangtong Zhang and Richard S Sutton · 2017
Cited alongside, same era.
Minimalistic gridworld environment for openai gym
Maxime Chevalier-Boisvert, Lucas Willems, and Suman Pal · 2018
Cited alongside, same era.
IMPALA: Scalable distributed deep-RL with importance weighted actor-learner architectures
Lasse Espeholt, Hubert Soyer, Remi Munos, Karen Simonyan, Vlad Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, et al · 2018
Cited alongside, same era.
Gan Memory with No Forgetting
Yulai Cong, Miaoyun Zhao, Jianqiao Li, Sijia Wang, and Lawrence Carin · 2020
Later among the works it cites.
Remind your neural network to prevent catastrophic forgetting
Tyler L Hayes, Kushal Kafle, Robik Shrestha, Manoj Acharya, and Christopher Kanan · 2020
Later among the works it cites.
Towards continual reinforcement learning: A review and perspectives
Khimya Khetarpal, Matthew Riemer, Irina Rish, and Doina Precup · 2020
Later among the works it cites.
Lifelong policy gradient learning of factored policies for faster training without forgetting
Jorge Mendez, Boyu Wang, and Eric Eaton · 2020
Later among the works it cites.
Ader: Adaptively distilled exemplar replay towards continual learning for session-based recommendation
Fei Mi, Xiaoyu Lin, and Boi Faltings · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Selective experience replay for lifelong learning
David Isele and Akansel Cosgun · 2018
Cited alongside, same era.
Progress & compress: A scalable framework for continual learning
Jonathan Schwarz, Wojciech Czarnecki, Jelena Luketina, Agnieszka Grabska-Barwinska, Yee Whye Teh, Razvan Pascanu, and Raia Hadsell · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Cited alongside, same era.
Generative replay with feedback connections as a general strategy for continual learning
Gido M van de Ven and Andreas S Tolias · 2018
Cited alongside, same era.
Empirical analysis of hidden technical debt patterns in machine learning software
Mohannad Alahdab and Gül Çalıklı · 2019
Cited alongside, same era.
Uncertainty-based modulation for lifelong learning
Andrew P Brna, Ryan C Brown, Patrick M Connolly, Stephen B Simons, Renee E Shimizu, and Mario Aguilar-Simon · 2019
Cited alongside, same era.
Sliced cramer synaptic consolidation for preserving deeply learned representations
Soheil Kolouri, Nicholas A Ketz, Andrea Soltoggio, and Praveen K Pilly · 2019
Cited alongside, same era.
Aswin Raghavan, Jesse Hostetler, Indranil Sur, Abrar Rahman, and Ajay Divakaran · 2020
Later among the works it cites.
Brain-inspired replay for continual learning with artificial neural networks
Gido M van de Ven, Hava T Siegelmann, and Andreas S Tolias · 2020
Later among the works it cites.
Pseudo-rehearsal: Achieving deep reinforcement learning without catastrophic forgetting
Craig Atkinson, Brendan McCane, Lech Szymanski, and Anthony Robins · 2021
Later among the works it cites.
Characterizing technical debt and antipatterns in ai-based systems: A systematic mapping study
Justus Bogner, Roberto Verdecchia, and Ilias Gerostathopoulos · 2021
Later among the works it cites.
Watch: Wasserstein change point detection for high-dimensional time series data
Kamil Faber, Roberto Corizzo, Bartlomiej Sniezynski, Michael Baron, and Nathalie Japkowicz · 2021
Later among the works it cites.
Self-supervised training enhances online continual learning
Jhair Gallardo, Tyler L Hayes, and Christopher Kanan · 2021
Later among the works it cites.
Replay in deep learning: Current approaches and missing biological elements
Tyler L Hayes, Giri P Krishnan, Maxim Bazhenov, Hava T Siegelmann, Terrence J Sejnowski, and Christopher Kanan · 2021
Later among the works it cites.
Continuous coordination as a realistic scenario for lifelong learning
Hadi Nekoei, Akilesh Badrinaaraayanan, Aaron Courville, and Sarath Chandar · 2021
Later among the works it cites.
Stable-baselines3: Reliable reinforcement learning implementations
Antonin Raffin, Ashley Hill, Adam Gleave, Anssi Kanervisto, Maximilian Ernestus, and Noah Dormann · 2021
Later among the works it cites.
Acae-remind for online continual learning with compressed feature replay
Kai Wang, Joost van de Weijer, and Luis Herranz · 2021
Later among the works it cites.
Alexander New, Megan Baker, Eric Nguyen, and Gautam Vallabha · 2022
Closest in time.