Fetching the paper…
Reading the bibliography…
Policies trained via Reinforcement Learning (RL) are often needlessly complex, making them difficult to analyse and interpret.
Empirical evaluation of the Tarantula automatic fault-localization technique
James A. Jones and Mary Jean Harrold · 2005
Earlier work this paper cites.
On the accuracy of spectrum-based fault localization
Rui Abreu, Peter Zoeteweij, and Arjan J.C. van Gemund · 2007
Earlier work this paper cites.
Automatic error detection techniques based on dynamic invariants
Alberto Gonzalez-Sanchez · 2007
Earlier work this paper cites.
Effective fault localization using code coverage
W. Eric Wong, Yu Qi, Lei Zhao, and Kai-Yuan Cai · 2007
Earlier work this paper cites.
A model for spectra-based software diagnosis
Lee Naish, Hua Jie Lee, and Kotagiri Ramamohanarao · 2011
Earlier work this paper cites.
Are automated debugging techniques actually helping programmers?
Chris Parnin and Alessandro Orso · 2011
Earlier work this paper cites.
Extended comprehensive study of association measures for fault localization
Lucia, David Lo, Lingxiao Jiang, Ferdian Thung, and Aditya Budi · 2014
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
A framework for automatic debugging of functional and degradation failures
Nuno Cardoso, Rui Abreu, Alexander Feldman, and Johan de Kleer · 2016
Earlier work this paper cites.
Why should I trust you? Explaining the predictions of any classifier
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin · 2016
Earlier work this paper cites.
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2016
Earlier work this paper cites.
Dueling network architectures for deep reinforcement learning
Ziyu Wang, Tom Schaul, Matteo Hessel, Hado Hasselt, Marc Lanctot, and Nando de Freitas · 2016
Cited alongside, same era.
A survey on software fault localization
W. Eric Wong, Ruizhi Gao, Yihao Li, Rui Abreu, and Franz Wotawa · 2016
Cited alongside, same era.
Automated debugging considered harmful: A user study revisiting the usefulness of spectra-based fault localization techniques with professionals using real bugs from large systems
Xin Xia, Lingfeng Bao, David Lo, and Shanping Li · 2016
Cited alongside, same era.
Graying the black box: Understanding DQNs
Tom Zahavy, Nir Ben-Zrihem, and Shie Mannor · 2016
Cited alongside, same era.
Improving robot controller transparency through autonomous policy explanation
Bradley Hayes and Julie A. Shah · 2017
Cited alongside, same era.
Programmatically interpretable reinforcement learning
Abhinav Verma, Vijayaraghavan Murali, Rishabh Singh, Pushmeet Kohli, and Swarat Chaudhuri · 2018
Later among the works it cites.
lcswillems/rl-starter-files, August 2020
Lucas Willems · 2018
Later among the works it cites.
DARPA’s explainable artificial intelligence program
David Gunning and David W. Aha · 2019
Later among the works it cites.
An Atari model zoo for analyzing, visualizing, and comparing deep reinforcement learning agents
Felipe Petroski Such, Vashisht Madhavan, Rosanne Liu, Rui Wang, Pablo Samuel Castro, Yulun Li, Jiale Zhi, Ludwig Schubert, Marc G. Bellemare, Jeff Clune, and Joel Lehman · 2019
Later among the works it cites.
Generation of policy-level explanations for reinforcement learning
Nicholay Topin and Manuela Veloso · 2019
Later among the works it cites.
The emerging landscape of explainable automated planning & decision making
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A unified approach to interpreting model predictions
Scott M. Lundberg and Su-In Lee · 2017
Cited alongside, same era.
Minimalistic gridworld environment for OpenAI Gym
Maxime Chevalier-Boisvert, Lucas Willems, and Suman Pal · 2018
Cited alongside, same era.
Rationalization: A neural machine translation approach to generating natural language explanations
Upol Ehsan, Brent Harrison, Larry Chan, and Mark O. Riedl · 2018
Cited alongside, same era.
Visualizing and understanding Atari agents
Samuel Greydanus, Anurag Koul, Jonathan Dodge, and Alan Fern · 2018
Cited alongside, same era.
Interpretable policies for reinforcement learning by genetic programming
Daniel Hein, Steffen Udluft, and Thomas A. Runkler · 2018
Cited alongside, same era.
Transparency and explanation in deep reinforcement learning neural networks
Rahul Iyer, Yuezhang Li, Huao Li, Michael Lewis, Ramitha Sundar, and Katia Sycara · 2018
Cited alongside, same era.
A practical evaluation of spectrum-based fault localization
Rui Abreu, Peter Zoeteweij, Rob Golsteijn, and Arjan J.C. van Gemund
Cited in the paper.
Tathagata Chakraborti, Sarath Sreedharan, and Subbarao Kambhampati · 2020
Closest in time.
g6ling/Reinforcement-Learning-Pytorch-Cartpole, July 2020
Cheol Kang · 2020
Closest in time.
PoPS: Policy pruning and shrinking for deep reinforcement learning
Dor Livne and Kobi Cohen · 2020
Closest in time.
TLdR: Policy summarization for factored SSP problems using temporal abstractions
Sarath Sreedharan, Siddharth Srivastava, and Subbarao Kambhampati · 2020
Closest in time.
Explaining image classifiers using statistical fault localization
Youcheng Sun, Hana Chockler, Xiaowei Huang, and Daniel Kroening · 2020
Closest in time.
Deep learning, transparency, and trust in human robot teamwork
Michael Lewis, Huao Li, and Katia Sycara · 2021
Closest in time.