Fetching the paper…
Reading the bibliography…
High-dimensional always-changing environments constitute a hard challenge for current reinforcement learning techniques.
1902
Earlier work this paper cites.
Lomonaco, V., Maltoni, D., Pellegrini, L.: Fine-Grained Continual Learning pp. 1–14 (2019),
1907
Earlier work this paper cites.
McCloskey, M., Cohen, N.J.: Catastrophic Interference in Connectionist Networks: The Sequential Learning Problem. Psychology of Learning and Motivation - Advances in Research and Theory
1989
Earlier work this paper cites.
Ring, M.: Continual Learning in Reinforcement Environments. Ph.D. thesis (1994)
1994
Earlier work this paper cites.
Robins, A.: Catastrophic Forgetting, Rehearsal and Pseudorehearsal. Connection Science
1995
Earlier work this paper cites.
Thrun, S., Mitchell, T.M.: Lifelong Robot Learning. The biology and technology of intelligent autonomous agents pp. 165—-196 (1995)
1995
Earlier work this paper cites.
French, R.M.: Catastrophic forgetting in connectionist networks. Trends in Cognitive Sciences
1999
Earlier work this paper cites.
2007
Earlier work this paper cites.
Clopath, C., Ziegler, L., Vasilaki, E., Büsing, L., Gerstner, W.: Tag-trigger-consolidation: A model of early and late long-term-potentiation and depression. PLoS Computational Biology
2008
Earlier work this paper cites.
Clopath, C.: Synaptic consolidation: An approach to long-term learning. Cognitive Neurodynamics
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
Mermillod, M., Bugaiska, A., Bonin, P.: The stability-plasticity dilemma: investigating the continuum from catastrophic forgetting to age-limited learning effects. Frontiers in psychology
2013
Earlier work this paper cites.
Seyed Hadi Mir Yazdi, Lashkari, Z.H.: Technical analysis of Forex by MACD Indicator. International Journal of Humanities and Management Sciences (IJHMS)
2013
Earlier work this paper cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., Hassabis, D.: Human-level control through deep reinforcement learning. Nature
2015
Earlier work this paper cites.
Peng, X., Sun, B., Ali, K., Saenko, K.: Learning Deep Object Detectors from 3D Models. In: 2015 IEEE International Conference on Computer Vision (ICCV). vol. 2015 Inter, pp. 1278–1286. IEEE (dec 2015). https://doi.org/10.1109/ICCV.2015.151,
2015
Earlier work this paper cites.
Beattie, C., Leibo, J.Z., Teplyashin, D., Ward, T., Wainwright, M., Lefrancq, A., Green, S., Sadik, A., Schrittwieser, J., Anderson, K., York, S., Cant, M., Cain, A., Bolton, A., Gaffney, S., King, H., Hassabis, D., Legg, S., Petersen, S.: DeepMind Lab pp. 1–11 (2016)
2016
Earlier work this paper cites.
Benna, M.K., Fusi, S.: Computational principles of synaptic memory consolidation. Nature Neuroscience
2016
Earlier work this paper cites.
Johnson, M., Hofmann, K., Hutton, T., Bignell, D.: The malmo platform for artificial intelligence experimentation. In: IJCAI International Joint Conference on Artificial Intelligence. vol. 2016-Janua, pp. 4246–4247 (2016)
2016
Cited alongside, same era.
2016
Cited alongside, same era.
Mnih, V., Badia, A.P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., Kavukcuoglu, K.: Asynchronous methods for deep reinforcement learning. American Journal of Health Behavior pp. 1928—-1937 (2016). https://doi.org/10.5993/AJHB.32.3.8
2016
Cited alongside, same era.
Vezhnevets, A.S., Osindero, S., Schaul, T., Heess, N., Jaderberg, M., Silver, D., Kavukcuoglu, K.: FeUdal Networks for Hierarchical Reinforcement Learning (2016)
2016
2018
Later among the works it cites.
2018
Later among the works it cites.
Lesort, T., Diaz-Rodriguez, N., Goudou, J.F., Filliat, D.: State representation learning for control: An overview. Neural Networks
2018
Later among the works it cites.
Machado, M.C., Bellemare, M.G., Talvitie, E., Veness, J., Hausknecht, M., Bowling, M.: Revisiting the arcade learning environment: Evaluation protocols and open problems for general agents. IJCAI International Joint Conference on Artificial Intelligence
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Arulkumaran, K., Deisenroth, M.P., Brundage, M., Bharath, A.A.: Deep Reinforcement Learning: A Brief Survey. IEEE Signal Processing Magazine
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Kansky, K., Silver, T., Miguel, E., Lou, X.: Schema Networks: Zero-shot Transfer with a Generative Causal Model of Intuitive Physics. International Conference on Machine Learning (2017)
2017
Cited alongside, same era.
Kirkpatrick, J., Pascanu, R., Rabinowitz, N., Veness, J., Desjardins, G., Rusu, A.A., Milan, K., Quan, J., Ramalho, T., Grabska-barwinska, A., Hassabis, D., Clopath, C., Kumaran, D., Hadsell, R.: Overcoming catastrophic forgetting in neural networks
2017
Cited alongside, same era.
Lomonaco, V., Maltoni, D.: CORe50: a New Dataset and Benchmark for Continuous Object Recognition. In: Levine, S., Vanhoucke, V., Goldberg, K. (eds.) Proceedings of the 1st Annual Conference on Robot Learning. Proceedings of Machine Learning Research, vol. 78, pp. 17–26. PMLR (2017),
2017
Cited alongside, same era.
Lopez-paz, D., Ranzato, M.: Gradient Episodic Memory for Continuum Learning. In: Advances in neural information processing systems (NIPS 2017) (2017),
2017
Cited alongside, same era.
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., Chen, Y., Lillicrap, T., Hui, F., Sifre, L., van den Driessche, G., Graepel, T., Hassabis, D.: Mastering the game of Go without human knowledge. Nature
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Later among the works it cites.
Riemer, M., Cases, I., Ajemian, R., Liu, M., Rish, I., Tu, Y., Tesauro, G.: Continual Learning by Maximizing Transfer and Minimizing Interference (NeurIPS), 1–24 (2018)
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
Wu, B.: Model Primitive Hierarchical Lifelong Reinforcement Learning (2018)
2018
Later among the works it cites.
Chevalier-Boisvert, M., Bahdanau, D., Lahlou, S., Willems, L., Saharia, C., Thien, H.N., Bengio, Y.: BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning. 7th International Conference on Learning Representations, (ICLR) pp. 1–18 (2019)
2019
Closest in time.
Liu, R., Zou, J.: The Effects of Memory Replay in Reinforcement Learning. 2018 56th Annual Allerton Conference on Communication, Control, and Computing, Allerton 2018 pp. 478–485 (2019)
2019
Closest in time.
Lomonaco, V.: Continual Learning with Deep Architectures. Phd thesis, University of Bologna (2019). https://doi.org/10.6092/unibo/amsdottorato/9073,
2019
Closest in time.
Maltoni, D., Lomonaco, V.: Continuous learning in single-incremental-task scenarios. Neural Networks
2019
Closest in time.
Parisi, G.I., Kemker, R., Part, J.L., Kanan, C., Wermter, S.: Continual lifelong learning with neural networks: A review. Neural Networks
2019
Closest in time.
Parisi, G.I., Lomonaco, V.: Online Continual Learning on Sequences, pp. 197–221. Springer International Publishing, Cham (2020)
2020
Closest in time.