Fetching the paper…
Reading the bibliography…
The Arcade Learning Environment (ALE) has become an essential benchmark for assessing the performance of reinforcement learning algorithms.
Juran on leadership for quality
J. M. Juran · 2003
Earlier work this paper cites.
Universal intelligence: A definition of machine intelligence
S. Legg and M. Hutter · 2007
Earlier work this paper cites.
Pearson correlation coefficient
J. Benesty, J. Chen, Y. Huang, and I. Cohen · 2009
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Earlier work this paper cites.
Compress and control
J. Veness, M. G. Bellemare, M. Hutter, A. Chua, and G. Desjardins · 2015
Earlier work this paper cites.
C. Beattie, J. Z. Leibo, D. Teplyashin, T. Ward, M. Wainwright, H. Küttler, A. Lefrancq, S. Green, V. Valdés, A. Sadik, et al · 2016
Earlier work this paper cites.
Unifying count-based exploration and intrinsic motivation
M. Bellemare, S. Srinivasan, G. Ostrovski, T. Schaul, D. Saxton, and R. Munos · 2016
Earlier work this paper cites.
Vizdoom: A doom-based ai research platform for visual reinforcement learning
M. Kempka, M. Wydmuch, G. Runc, J. Toczek, and W. Jaśkowski · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Playing hard exploration games by watching youtube
Y. Aytar, T. Pfaff, D. Budden, T. Paine, Z. Wang, and N. De Freitas · 2018
Cited alongside, same era.
Implicit quantile networks for distributional reinforcement learning
W. Dabney, G. Ostrovski, D. Silver, and R. Munos · 2018
Cited alongside, same era.
The reactor: A fast and sample-efficient actor-critic agent for reinforcement learning
A. Gruslys, W. Dabney, M. G. Azar, B. Piot, M. Bellemare, and R. Munos · 2018
Cited alongside, same era.
Rainbow: Combining improvements in deep reinforcement learning
Minatar: An atari-inspired testbed for thorough and reproducible reinforcement learning experiments
K. Young and T. Tian · 2019
Later among the works it cites.
Agent57: Outperforming the atari human benchmark
A. P. Badia, B. Piot, S. Kapturowski, P. Sprechmann, A. Vitvitskyi, Z. D. Guo, and C. Blundell · 2020
Later among the works it cites.
Leveraging procedural generation to benchmark reinforcement learning
K. Cobbe, C. Hesse, J. Hilton, and J. Schulman · 2020
Later among the works it cites.
Accelerating reinforcement learning through gpu atari emulation
S. Dalton et al · 2020
Later among the works it cites.
Mastering atari, go, chess and shogi by planning with a learned model
J. Schrittwieser, I. Antonoglou, T. Hubert, K. Simonyan, L. Sifre, S. Schmitt, A. Guez, E. Lockhart, D. Hassabis, T. Graepel, et al · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Hessel, J. Modayil, H. Van Hasselt, T. Schaul, G. Ostrovski, W. Dabney, D. Horgan, B. Piot, M. Azar, and D. Silver · 2018
Cited alongside, same era.
Distributed prioritized experience replay
D. Horgan, J. Quan, D. Budden, G. Barth-Maron, M. Hessel, H. van Hasselt, and D. Silver · 2018
Cited alongside, same era.
Recurrent experience replay in distributed reinforcement learning
S. Kapturowski, G. Ostrovski, J. Quan, R. Munos, and W. Dabney · 2018
Cited alongside, same era.
Gotta learn fast: A new benchmark for generalization in rl
A. Nichol, V. Pfau, C. Hesse, O. Klimov, and J. Schulman · 2018
Cited alongside, same era.
Energy and policy considerations for deep learning in nlp
E. Strubell, A. Ganesh, and A. McCallum · 2019
Cited alongside, same era.
Revisiting rainbow: Promoting more insightful and inclusive deep reinforcement learning research
J. S. O. Ceron and P. S. Castro · 2021
Later among the works it cites.
First return, then explore
A. Ecoffet, J. Huizinga, J. Lehman, K. O. Stanley, and J. Clune · 2021
Later among the works it cites.
Carbon emissions and large neural network training
D. Patterson, J. Gonzalez, Q. Le, C. Liang, L.-M. Munguia, D. Rothchild, D. So, M. Texier, and J. Dean · 2021
Later among the works it cites.
Conjugated discrete distributions for distributional reinforcement learning
B. Lindenberg, J. Nordqvist, and K.-O. Lindahl · 2022
Closest in time.