Fetching the paper…
Reading the bibliography…
The Arcade Learning Environment (ALE) is a popular platform for evaluating reinforcement learning agents.
“Metatrace: Online Step-size Tuning by Meta-gradient Descent for Reinforcement Learning Control”
Kenny Young, Baoxiang Wang and Matthew Taylor · 1903
Earlier work this paper cites.
“Model-free reinforcement learning with continuous action in practice”
Thomas Degris, Patrick Pilarski and Richard Sutton · 2012
Earlier work this paper cites.
“Lecture 6.5-RMSProp: Divide the gradient by a running average of its recent magnitude”
Tijmen Tieleman and Geoffrey Hinton · 2012
Earlier work this paper cites.
“Mujoco: A physics engine for model-based control”
Emanuel Todorov, Tom Erez and Yuval Tassa · 2012
Earlier work this paper cites.
“The Arcade Learning Environment: An Evaluation Platform for General Agents”
M.. Bellemare, Y. Naddaf, J. Veness and M. Bowling · 2013
Earlier work this paper cites.
“Adam: A method for stochastic optimization”
Diederik Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
“Human-level control through deep reinforcement learning”
Volodymyr Mnih et al · 2015
Cited alongside, same era.
“Mastering the game of Go with deep neural networks and tree search”
David Silver et al · 2016
Cited alongside, same era.
“PyGame Learning Environment”
Norman Tasfi · 2016
Cited alongside, same era.
Marlos. Machado et al · 2017
Cited alongside, same era.
“Self-correcting models for model-based reinforcement learning”
Erik Talvitie · 2017
Cited alongside, same era.
“StarCraft II: A new challenge for reinforcement learning”
Oriol Vinyals et al · 2017
Later among the works it cites.
“Sigmoid-weighted linear units for neural network function approximation in reinforcement learning”
Stefan Elfwing, Eiji Uchibe and Kenji Doya · 2018
Later among the works it cites.
“Deep reinforcement learning that matters”
Peter Henderson et al · 2018
Later among the works it cites.
“Reinforcement learning: An introduction”
Richard Sutton and Andrew Barto · 2018
Later among the works it cites.
Yuval Tassa et al · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…