Fetching the paper…
Reading the bibliography…
Saliency maps are frequently used to support explanations of the behavior of deep reinforcement learning (RL) agents.
Spurious correlation: A causal interpretation
Herbert A. Simon · 1954
Earlier work this paper cites.
The logic of scientific discovery
Karl Popper · 1959
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Causality: models, reasoning and inference , volume 29
Judea Pearl · 2000
Earlier work this paper cites.
Learning causal laws
Joshua B Tenenbaum and Sourabh Niyogi · 2003
Earlier work this paper cites.
The similarity of causal inference in experimental and non‐experimental studies
Richard Scheines · 2005
Earlier work this paper cites.
Template Matching Techniques in Computer Vision: Theory and Practice
Roberto Brunelli · 2009
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman · 2014
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Matthew D Zeiler and Rob Fergus · 2014
Earlier work this paper cites.
Deep apprenticeship learning for playing video games
Miroslav Bogdanovic, Dejan Markovikj, Misha Denil, and Nando De Freitas · 2015
Earlier work this paper cites.
Visual causal feature learning
Krzysztof Chalupka, Pietro Perona, and Frederick Eberhardt · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Striving for simplicity: The all convolutional net
Jost Tobias Springenberg, Alexey Dosovitskiy, Thomas Brox, and Martin Riedmiller · 2015
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Why should I trust you?: Explaining the predictions of any classifier
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin · 2016
Earlier work this paper cites.
Dueling network architectures for deep reinforcement learning
Ziyu Wang, Tom Schaul, Matteo Hessel, Hado Van Hasselt, Marc Lanctot, and Nando De Freitas · 2016
Earlier work this paper cites.
Graying the black box: Understanding dqns
Tom Zahavy, Nir Ben-Zrihem, and Shie Mannor · 2016
Cited alongside, same era.
Real time image saliency for black box classifiers
Piotr Dabkowski and Yarin Gal · 2017
Cited alongside, same era.
How people explain action (and autonomous intelligent systems should too)
Maartje M. A. de Graaf and Bertram F. Malle · 2017
Cited alongside, same era.
OpenAI Baselines
Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, and Yuhuai Wu · 2017
Cited alongside, same era.
Interpretable explanations of black boxes by meaningful perturbation
Ruth C Fong and Andrea Vedaldi · 2017
Cited alongside, same era.
Visualizing and understanding atari agents
Sam Greydanus, Anurag Koul, Jonathan Dodge, and Alan Fern · 2017
Cited alongside, same era.
Noise-adding methods of saliency map as series of higher order partial derivative
Junghoon Seo, Jeongyeol Choe, Jamyoung Koo, Seunghyeon Jeon, Beomsu Kim, and Taegyun Jeon · 2018
Later among the works it cites.
Transparency in Deep Reinforcement Learning Networks
Ramitha Sundar · 2018
Later among the works it cites.
Advantage actor-critic methods for carracing
Douwe van der Wal, Bachelor Opleiding Kunstmatige Intelligentie, and Wenling Shang · 2018
Later among the works it cites.
Programmatically interpretable reinforcement learning
Abhinav Verma, Vijayaraghavan Murali, Rishabh Singh, Pushmeet Kohli, and Swarat Chaudhuri · 2018
Later among the works it cites.
Learn to interpret atari agents
Zhao Yang, Song Bai, Li Zhang, and Philip HS Torr · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Evaluating the visualization of what a deep neural network has learned
Wojciech Samek, Alexander Binder, Grégoire Montavon, Sebastian Lapuschkin, and Klaus-Robert Müller · 2017
Cited alongside, same era.
Grad-cam: Visual explanations from deep networks via gradient-based localization
Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, Dhruv Batra, et al · 2017
Cited alongside, same era.
Learning important features through propagating activation differences
Avanti Shrikumar, Peyton Greenside, and Anshul Kundaje · 2017
Cited alongside, same era.
Smoothgrad: removing noise by adding noise
Daniel Smilkov, Nikhil Thorat, Been Kim, Fernanda Viégas, and Martin Wattenberg · 2017
Cited alongside, same era.
Sanity checks for saliency maps
Julius Adebayo, Justin Gilmer, Michael Muelly, Ian Goodfellow, Moritz Hardt, and Been Kim · 2018
Cited alongside, same era.
Investigating human priors for playing video games
Rachit Dubey, Pulkit Agrawal, Deepak Pathak, Thomas L Griffiths, and Alexei A Efros · 2018
Cited alongside, same era.
Jianming Zhang, Sarah Adel Bargal, Zhe Lin, Jonathan Brandt, Xiaohui Shen, and Stan Sclaroff · 2018
Later among the works it cites.
Towards better interpretability in deep q-networks
Raghuram Mandyam Annasamy and Katia Sycara · 2019
Closest in time.
Counterfactuals in explainable artificial intelligence (XAI): Evidence from human reasoning
Ruth M. J. Byrne · 2019
Closest in time.
Evaluating feature importance estimates
Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans, and Been Kim · 2019
Closest in time.
Explainable reinforcement learning via reward decomposition
Zoe Juozapaitis, Anurag Koul, Alan Fern, Martin Erwig, and Finale Doshi-Velez · 2019
Closest in time.
The (un) reliability of saliency methods
Pieter-Jan Kindermans, Sara Hooker, Julius Adebayo, Maximilian Alber, Kristof T Schütt, Sven Dähne, Dumitru Erhan, and Been Kim · 2019
Closest in time.
Explaining explanations in ai
Brent Mittelstadt, Chris Russell, and Sandra Wachter · 2019
Closest in time.
Towards interpretable reinforcement learning using attention augmented agents
Alex Mott, Daniel Zoran, Mike Chrzanowski, Daan Wierstra, and Danilo J Rezende · 2019
Closest in time.
Free-lunch saliency via attention in atari agents
Dmitry Nikulin, Anastasia Ianina, Vladimir Aliev, and Sergey Nikolenko · 2019
Closest in time.
Counterfactual states for atari agents via generative deep learning
Matthew L Olson, Lawrence Neal, Fuxin Li, and Weng-Keen Wong · 2019
Closest in time.
Toybox: A suite of environments for experimental evaluation of deep reinforcement learning, 2019
Emma Tosch, Kaleigh Clary, John Foley, and David Jensen · 2019
Closest in time.