Maps of bounded rationality: Psychology for behavioral economics †
Daniel Kahneman · 2003
Earlier work this paper cites.
Uncertainty-based competition between prefrontal and dorsolateral striatal systems for behavioral control
Nathaniel D Daw, Yael Niv, and Peter Dayan · 2005
Earlier work this paper cites.
An empirical investigation of catastrophic forgetting in gradient-based neural networks, 2013
Ian J. Goodfellow, Mehdi Mirza, Da Xiao, Aaron Courville, and Yoshua Bengio · 2013
Earlier work this paper cites.
Guided policy search
Sergey Levine and Vladlen Koltun · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Original
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller · 2013
Earlier work this paper cites.
Ella: An efficient lifelong learning algorithm
Paul Ruvolo and Eric Eaton · 2013
Earlier work this paper cites.
Heuristic and analytic processes in reasoning: An event-related potential study of belief bias
Adrian Banks and Christopher Hope · 2014
Earlier work this paper cites.
Model-based and model-free pavlovian reward learning: Revaluation, revision, and revelation
Peter Dayan and Kent C Berridge · 2014
Earlier work this paper cites.
Non-stationary stochastic optimization
Omar Besbes, Yonatan Gur, and Assaf Zeevi · 2015
Earlier work this paper cites.
Interactive control of diverse complex characters with neural networks
Igor Mordatch, Kendall Lowrey, Galen Andrew, Zoran Popovic, and Emanuel V Todorov · 2015
Earlier work this paper cites.
Model predictive path integral control using covariance variable importance sampling
Original
Grady Williams, Andrew Aldrich, and Evangelos A. Theodorou · 2015
Earlier work this paper cites.
Regret bounds for lifelong learning, 2016
Pierre Alquier, The Tien Mai, and Massimiliano Pontil · 2016
Earlier work this paper cites.
Openai gym
Original
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Deep exploration via bootstrapped DQN
Original
Ian Osband, Charles Blundell, Alexander Pritzel, and Benjamin Van Roy · 2016
Earlier work this paper cites.