Fetching the paper…
Reading the bibliography…
Episodic control has been proposed as a third approach to reinforcement learning, besides model-free and model-based control, by analogy with the three types of human memory.
Dyna, an integrated architecture for learning, planning, and reacting
Sutton, Richard S · 1991
Earlier work this paper cites.
Prioritized sweeping: Reinforcement learning with less data and less time
Moore, Andrew W. and Atkeson, Christopher G · 1993
Earlier work this paper cites.
Efficient learning and planning within the dyna framework
Peng, Jing and Williams, R. J · 1993
Earlier work this paper cites.
Reverse replay of behavioural sequences in hippocampal place cells during the awake state
Foster, David J. and Wilson, Matthew A · 2006
Earlier work this paper cites.
Episodic memory
Clayton, Nicola S., Salwiczek, Lucie H., and Dickinson, Anthony · 2007
Earlier work this paper cites.
Hippocampal contributions to control: The third way
Lengyel, Máté and Dayan, Peter · 2008
Cited alongside, same era.
Hippocampal replay is not a simple function of experience
Gupta, Anoopum S., van der Meer, Matthijs A.A., Touretzky, David S., and Redish, A. David · 2010
Cited alongside, same era.
Planning by prioritized sweeping with small backups
Seijen, Harm Van and Sutton, Rich · 2013
Cited alongside, same era.
Model-Free Episodic Control
Blundell, C., Uria, B., Pritzel, A., Li, Y., Ruderman, A., Leibo, J. Z, Rae, J., Wierstra, D., and Hassabis, D · 2016
Cited alongside, same era.
The Concrete Distribution: A Continuous Relaxation of Discrete Random Variables
Maddison, C. J., Mnih, A., and Whye Teh, Y · 2016
Cited alongside, same era.
Neural Episodic Control
Pritzel, A., Uria, B., Srinivasan, S., Puigdomènech, A., Vinyals, O., Hassabis, D., Wierstra, D., and Blundell, C · 2017
Closest in time.
Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Silver, D., Hubert, T., Schrittwieser, J., Antonoglou, I., Lai, M., Guez, A., Lanctot, M., Sifre, L., Kumaran, D., Graepel, T., Lillicrap, T., Simonyan, K., and Hassabis, D · 2017
Closest in time.
Reinforcement Learning: An Introduction
Sutton, Richard S. and Barto, Andrew G · 2018
Closest in time.
Reinforcement learning and episodic memory in humans and animals: An integrative framework
Gershman, Samuel J. and Daw, Nathaniel D · 2085
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…