Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) has achieved significant breakthroughs in various tasks.
STRIPS: A New Approach to the Application of Theorem Proving to Problem Solving
Fikes, R. E. and Nilsson, N. J · 1971
Earlier work this paper cites.
Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning
Willia, R. J · 1992
Earlier work this paper cites.
Introduction to Reinforcement Learning
Sutton, R. S. and Barto, A. G · 1998
Earlier work this paper cites.
Symbolic Dynamic Programming for First-order MDPs
Boutilier, C., Reiter, R., and Price, B · 2001
Earlier work this paper cites.
Relational reinforcement learning
Džeroski, S., De Raedt, L., and Driessens, K · 2001
Earlier work this paper cites.
Inductive Policy Selection for First-order MDPs
Yoon, S., Fern, A., and Givan, R · 2002
Earlier work this paper cites.
Relational instance based regression for relational reinforcement learning
Driessens, K. and Ramon, J · 2003
Earlier work this paper cites.
Generalizing Plans to New Environments in Relational MDPs
Guestrin, C., Koller, D., Gearhart, C., and Kanodia, N · 2003
Earlier work this paper cites.
Integrating guidance into relational reinforcement learning
Driessens, K. and Džeroski, S · 2004
Earlier work this paper cites.
Introduction to Statistical Relational Learning
Getoor, L. and Taskar, B · 2007
Cited alongside, same era.
Gradient-based relational reinforcement learning of temporally extended policies
Gretton, C · 2007
Cited alongside, same era.
Rectified linear units improve restricted boltzmann machines
Nair, V. and Hinton, G. E · 2010
Cited alongside, same era.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, N., Hinton, G., Krizhevsky, A., Sutskever, I., and Salakhutdinov, R · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., and Hassabis, D · 2015
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Later among the works it cites.
Mastering the game of Go without human knowledge
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., Chen, Y., Lillicrap, T., Hui, F., Sifre, L., Van Den Driessche, G., Graepel, T., and Hassabis, D · 2017
Later among the works it cites.
Mutual alignment transfer learning
Wulfmeier, M., Posner, I., and Abbeel, P · 2017
Later among the works it cites.
Quantifying the reality gap in robotic manipulation tasks
Collins, J., Howard, D., and Leitner, J · 2018
Later among the works it cites.
Learning Explanatory Rules from Noisy Data
Evans, R. and Grefenstette, E · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mnih, V., Badia, A. P., Mirza, M., Graves, A., Harley, T., Lillicrap, T. P., Silver, D., and Kavukcuoglu, K · 2016
Cited alongside, same era.
TensorLog: Deep Learning Meets Probabilistic DBs
Cohen, W. W., Yang, F., and Mazaitis, K. R · 2017
Cited alongside, same era.
Towards a rigorous science of interpretable machine learning
Doshi-Velez, F. and Kim, B · 2017
Cited alongside, same era.
End-to-end differentiable proving
Rocktäschel, T. and Riedel, S · 2017
Cited alongside, same era.
Methods for interpreting and understanding deep neural networks
Montavon, G., Samek, W., and Müller, K.-R
Cited in the paper.
Trust region policy optimization
Schulman, J., Levine, S., Moritz, P., Jordan, M., and Abbeel, P
Cited in the paper.
High-dimensional continuous control using generalized advantage estimation
Schulman, J., Moritz, P., Levine, S., Jordan, M. I., and Abbeel, P
Cited in the paper.
Nervenet: Learning structured policy with graph neural networks
Wang, T., Liao, R., Ba, J., and Fidler, S · 2018
Later among the works it cites.
Relational Deep Reinforcement Learning
Zambaldi, V., Raposo, D., Santoro, A., Bapst, V., Li, Y., Babuschkin, I., Tuyls, K., Reichert, D., Lillicrap, T., Lockhart, E., Shanahan, M., Langston, V., Pascanu, R., Botvinick, M., Vinyals, O., and Battaglia, P · 2018
Later among the works it cites.
A Study on Overfitting in Deep Reinforcement Learning
Zhang, C., Vinyals, O., Munos, R., and Bengio, S · 2018
Later among the works it cites.
Neural logic machines
Dong, H., Mao, J., Lin, T., Wang, C., Li, L., and Zhou, D · 2019
Closest in time.