Fetching the paper…
Reading the bibliography…
Continual Reinforcement Learning (CRL) is a challenging setting where an agent learns to interact with an environment that is constantly changing over time (the stream of experiences).
Robins, A.V.: Catastrophic forgetting, rehearsal and pseudorehearsal. Connect. Sci. 7
1995
Earlier work this paper cites.
Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., Fei-Fei, L.: Imagenet: A large-scale hierarchical image database. In: 2009 IEEE conference on computer vision and pattern recognition. pp. 248–255. Ieee (2009)
2009
Earlier work this paper cites.
LeCun, Y., Cortes, C.: MNIST handwritten digit database (2010), http://yann.lecun.com/exdb/mnist/
2010
Earlier work this paper cites.
2012
Earlier work this paper cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Graves, A., Antonoglou, I., Wierstra, D., Riedmiller, M.: Playing atari with deep reinforcement learning (2013)
2013
Earlier work this paper cites.
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., , et al.: TensorFlow: Large-scale machine learning on heterogeneous systems (2015), https://www.tensorflow.org/ , software available from tensorflow.org
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., et al.: Human-level control through deep reinforcement learning. Nature 518
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Cited alongside, same era.
Plappert, M.: keras-rl. https://github.com/keras-rl/keras-rl (2016)
2016
Cited alongside, same era.
2017
Cited alongside, same era.
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., Klimov, O.: Proximal policy optimization algorithms (2017)
2017
Cited alongside, same era.
Schwarz, J., Altman, D., Dudzik, A., Vinyals, O., Teh, Y.W., and, R.P.: Towards a natural benchmark for continual learning (2018), https://marcpickett.com/cl2018/CL-2018_paper_48.pdf
2018
Later among the works it cites.
Sutton, R.S., Barto, A.G.: Reinforcement Learning: An Introduction. The MIT Press, second edn. (2018), http://incompleteideas.net/book/the-book-2nd.html
2018
Later among the works it cites.
Lesort, T., Lomonaco, V., Stoian, A., Maltoni, D., Filliat, D., Díaz-Rodríguez, N.: Continual learning for robotics: Definition, framework, learning strategies, opportunities and challenges. Information Fusion 58
2019
Later among the works it cites.
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., et al.: Pytorch: An imperative style, high-performance deep learning library. In: Wallach, H., Larochelle, H., Beygelzimer, A., d'Alché-Buc, F., Fox, E., Garnett, R. (eds.) Advances in Neural Information Processing Systems 32, pp. 8024–8035. Curran Associates, Inc. (2019), http://papers.neurips.cc/paper/9015-pytorch-an-imperative-style-high-performance-deep-learning-library.pdf
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., et al.: Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures (2018)
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Hook, D.W., Porter, S.J., Herzog, C.: Dimensions: Building context for search and evaluation. Frontiers in Research Metrics and Analytics 3
2018
Cited alongside, same era.
Isele, D., Cosgun, A.: Selective experience replay for lifelong learning. CoRR abs/1802.10269
2018
Cited alongside, same era.
Moritz, P., Nishihara, R., Wang, S., Tumanov, A., Liaw, R., Liang, E., Elibol, M., Yang, Z., Paul, W., Jordan, M.I., Stoica, I.: Ray: A distributed framework for emerging ai applications (2018)
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Krizhevsky, A., Nair, V., Hinton, G.: Cifar-10 (canadian institute for advanced research) http://www.cs.toronto.edu/~kriz/cifar.html
Cited in the paper.
2019
Later among the works it cites.
Raffin, A., Hill, A., Ernestus, M., Gleave, A., Kanervisto, A., Dormann, N.: Stable baselines3. https://github.com/DLR-RM/stable-baselines3 (2019)
2019
Later among the works it cites.
Savva, M., Kadian, A., Maksymets, O., Zhao, Y., Wijmans, E., Jain, B., et al.: Habitat: A platform for embodied ai research (2019)
2019
Later among the works it cites.
2021
Later among the works it cites.
Lomonaco, V., Pellegrini, L., Cossu, A., Carta, A., Graffieti, G., Hayes, T.L., et al.: Avalanche: an end-to-end library for continual learning (2021)
2021
Later among the works it cites.
2021
Later among the works it cites.