Fetching the paper…
Reading the bibliography…
Efficient planning plays a crucial role in model-based reinforcement learning.
Learning to predict by the methods of temporal differences
Sutton, R. S. (1988) · 1988
Earlier work this paper cites.
Watkins, C. (1989) · 1989
Earlier work this paper cites.
Self-improving reactive agents based on reinforcement learning, planning and teaching
Lin, L.J. (1992) · 1992
Earlier work this paper cites.
Prioritized sweeping: Reinforcement learning with less data and less real time
Moore, A., Atkeson, C. (1993) · 1993
Cited alongside, same era.
Efficient learning and planning within the dyna framework
Peng, J., Williams, R. J. (1993) · 1993
Cited alongside, same era.
On step-size and bias in temporal-difference learning
Sutton, R. S., Singh, S. P. (1994) · 1994
Cited alongside, same era.
Reinforcement learning: A survey
Kaelbling, L. P., Littman, M. L., Moore, A. P. (1996) · 1996
Later among the works it cites.
Reinforcement Learning: An Introduction
Sutton, R. S., Barto, A. G. (1998) · 1998
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…