Fetching the paper…
Reading the bibliography…
In this paper, we investigate the problem of overfitting in deep reinforcement learning.
On the shortest spanning subtree of a graph and the traveling salesman problem
J. B. Kruskal · 1956
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling · 2012
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. A. Riedmiller · 2013
Earlier work this paper cites.
The impact of determinism on learning atari 2600 games
M. J. Hausknecht and P. Stone · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. A. Riedmiller, A. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
Earlier work this paper cites.
Improved regularization of convolutional neural networks with cutout
T. Devries and G. W. Taylor · 2017
Cited alongside, same era.
Openai baselines
P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, Y. Wu, and P. Zhokhov · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Cited alongside, same era.
Domain randomization for transferring deep neural networks from simulation to the real world
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel · 2017
Cited alongside, same era.
Autoaugment: Learning augmentation policies from data
E. D. Cubuk, B. Zoph, D. Mané, V. Vasudevan, and Q. V. Le · 2018
Generalization and regularization in DQN
J. Farebrother, M. C. Machado, and M. Bowling · 2018
Closest in time.
Illuminating generalization in deep reinforcement learning through procedural level generation
N. Justesen, R. R. Torrado, P. Bontrager, A. Khalifa, J. Togelius, and S. Risi · 2018
Closest in time.
Towards understanding regularization in batch normalization
P. Luo, X. Wang, W. Shao, and Z. Peng · 2018
Closest in time.
Revisiting the arcade learning environment: Evaluation protocols and open problems for general agents
M. C. Machado, M. G. Bellemare, E. Talvitie, J. Veness, M. J. Hausknecht, and M. Bowling · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
IMPALA: scalable distributed deep-rl with importance weighted actor-learner architectures
L. Espeholt, H. Soyer, R. Munos, K. Simonyan, V. Mnih, T. Ward, Y. Doron, V. Firoiu, T. Harley, I. Dunning, S. Legg, and K. Kavukcuoglu · 2018
Cited alongside, same era.
A dissection of overfitting and generalization in continuous reinforcement learning
A. Zhang, N. Ballas, and J. Pineau
Cited in the paper.
A study on overfitting in deep reinforcement learning
C. Zhang, O. Vinyals, R. Munos, and S. Bengio
Cited in the paper.
A. Nichol, V. Pfau, C. Hesse, O. Klimov, and J. Schulman · 2018
Closest in time.
Assessing generalization in deep reinforcement learning
C. Packer, K. Gao, J. Kos, P. Krähenbühl, V. Koltun, and D. Song · 2018
Closest in time.