Fetching the paper…
Reading the bibliography…
The premise of the Multi-disciplinary Conference on Reinforcement Learning and Decision Making is that multiple disciplines share an interest in goal-directed decision making over time.
The Nature of Explanation
Craik, K. J. W. (1943) · 1943
Earlier work this paper cites.
Tolman, E. C. (1948). Cognitive maps in rats and men. Psychological review, 55 (4), 189
1948
Earlier work this paper cites.
Hochreiter, S., Schmidhuber, J. (1997). Long short-term memory. Neural computation, 9 (8), 1735–1780
1997
Earlier work this paper cites.
A neural substrate of prediction and reward
Schultz, W., Dayan, P., Montague, P. R. (1997) · 1997
Cited alongside, same era.
Littman, M. L., Sutton, R. S., Singh, S. (2002). Predictive representations of state. In Advances in Neural Information Processing Systems 14 , pp. 1555–1561. MIT Press, Cambridge, MA
2002
Cited alongside, same era.
Glimcher, P. W. (2011). Understanding dopamine and reinforcement learning: The dopamine reward prediction error hypothesis. Proceedings of the National Academy of Sciences, 108 (Supplement 3), 15647–15654
2011
Cited alongside, same era.
Kahneman, D. (2011). Thinking, Fast and Slow . Macmillan
2011
Later among the works it cites.
Schrittwieser, J., Antonoglou, I., Hubert, T., Simonyan, K., Sifre, L., Schmitt, S., … Silver, D. (2020). Mastering atari, go, chess and shogi by planning with a learned model. Nature, 588 (7839), 604–609
2020
Later among the works it cites.
2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…