Fetching the paper…
Reading the bibliography…
Recent progress in deep reinforcement learning (DRL) can be largely attributed to the use of neural networks.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Self-supervised discovering of causal features: Towards interpretable reinforcement learning
Wenjie Shi, Zhuoyuan Wang, Shiji Song, and Gao Huang · 2003
Earlier work this paper cites.
Relational retrieval using a combination of path-constrained random walks
Ni Lao and William W Cohen · 2010
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Vinod Nair and Geoffrey E Hinton · 2010
Earlier work this paper cites.
Efficient and expressive knowledge base completion using subgraph feature extraction
Matt Gardner and Tom Mitchell · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2015
Earlier work this paper cites.
High-dimensional continuous control using generalized advantage estimation
John Schulman, Philipp Moritz, Sergey Levine, Michael Jordan, and Pieter Abbeel · 2015
Earlier work this paper cites.
Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation
Tejas D. Kulkarni, Karthik Narasimhan, Ardavan Saeedi, and Josh Tenenbaum · 2016
Earlier work this paper cites.
The mythos of model interpretability
Zachary C Lipton · 2016
Earlier work this paper cites.
Deep reinforcement learning with double q-learning
Hado Van Hasselt, Arthur Guez, and David Silver · 2016
Earlier work this paper cites.
Graying the black box: Understanding dqns
Tom Zahavy, Nir Ben-Zrihem, and Shie Mannor · 2016
Earlier work this paper cites.
Chains of reasoning over entities, relations, and text using recurrent neural networks
Rajarshi Das, Arvind Neelakantan, David Belanger, and Andrew Mccallum · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Differentiable learning of logical rules for knowledge base reasoning
Fan Yang, Zhilin Yang, and William W Cohen · 2017
Cited alongside, same era.
Explain your move: Understanding agent actions using focused feature saliency
Piyush Gupta, Nikaash Puri, Sukriti Verma, Dhruv Kayastha, Shripad Deshmukh, Balaji Krishnamurthy, and Sameer Singh · 2019
Later among the works it cites.
Neural logic reinforcement learning
Zhengyao Jiang and Shan Luo · 2019
Later among the works it cites.
Explainable reinforcement learning via reward decomposition
Zoe Juozapaitis, Anurag Koul, Alan Fern, Martin Erwig, and Finale Doshi-Velez · 2019
Later among the works it cites.
Sdrl: Interpretable and data-efficient deep reinforcement learning leveraging symbolic planning
Daoming Lyu, Fangkai Yang, Bo Liu, and Steven Gustafson · 2019
Later among the works it cites.
Generation of policy-level explanations for reinforcement learning
Nicholay Topin and Manuela Veloso · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Verifiable reinforcement learning via policy extraction
Osbert Bastani, Yewen Pu, and Armando Solar-Lezama · 2018
Cited alongside, same era.
Learning explanatory rules from noisy data
R Evans and E Grefenstette · 2018
Cited alongside, same era.
Visualizing and understanding atari agents
Sam Greydanus, Anurag Koul, Jonathan Dodge, and Alan Fern · 2018
Cited alongside, same era.
Programmatically interpretable reinforcement learning
Abhinav Verma, Vijayaraghavan Murali, Rishabh Singh, Pushmeet Kohli, and Swarat Chaudhuri · 2018
Cited alongside, same era.
Relational deep reinforcement learning
Vinicius Zambaldi, David Raposo, Adam Santoro, Victor Bapst, Yujia Li, Igor Babuschkin, Karl Tuyls, David Reichert, Timothy Lillicrap, Edward Lockhart, et al · 2018
Cited alongside, same era.
Distilling deep reinforcement learning policies in soft decision trees
Youri Coppens, Kyriakos Efthymiadis, Tom Lenaerts, and Ann Nowé · 2019
Cited alongside, same era.
Neural logic machines
Honghua Dong, Jiayuan Mao, Tian Lin, Chong Wang, Lihong Li, and Dengyong Zhou · 2019
Cited alongside, same era.
A Verma · 2019
Later among the works it cites.
Turning 30: New ideas in inductive logic programming
Andrew Cropper, Sebastijan Dumančić, and Stephen H Muggleton · 2020
Later among the works it cites.
Explainable reinforcement learning through a causal lens
Prashan Madumal, Tim Miller, Liz Sonenberg, and Frank Vetere · 2020
Later among the works it cites.
Ali Payani and Faramarz Fekri · 2020
Later among the works it cites.
Interestingness elements for explainable reinforcement learning: Understanding agents’ capabilities and limitations
Pedro Sequeira and Melinda Gervasio · 2020
Later among the works it cites.
Learn to explain efficiently via neural logic inductive learning
Yuan Yang and Le Song · 2020
Later among the works it cites.