Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (RL) has recently led to many breakthroughs on a range of complex control tasks.
T. Zahavy, N. Ben-Zrihem, and S. Mannor, “Graying the black box: Understanding dqns,” in International Conference on Machine Learning , 2016, pp. 1899–1908
1908
Earlier work this paper cites.
G. Lin, A. Milan, C. Shen, and I. Reid, “Refinenet: Multi-path refinement networks for high-resolution semantic segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 1925–1934
1934
Earlier work this paper cites.
L. S. Shapley, “A value for n-person games,” Contributions to the Theory of Games , vol. 2, no. 28, pp. 307–317, 1953
1953
Earlier work this paper cites.
Y. Engel and S. Mannor, “Learning embedded maps of markov processes,” in International Conference on Machine Learning , 2001
2001
Earlier work this paper cites.
Z. Wang, T. Schaul, M. Hessel, H. Van Hasselt, M. Lanctot, and N. De Freitas, “Dueling network architectures for deep reinforcement learning,” in International Conference on Machine Learning , vol. 48, 2016, pp. 1995–2003
2003
Earlier work this paper cites.
2003
Earlier work this paper cites.
F. Elizalde, L. E. Sucar, M. Luque, J. Diez, and A. Reyes, “Policy explanation in factored markov decision processes,” in Proceedings of the 4th European Workshop on Probabilistic Graphical Models (PGM 2008) , 2008, pp. 97–104
2008
Earlier work this paper cites.
L. v. d. Maaten and G. Hinton, “Visualizing data using t-sne,” Journal of Machine Learning Research , vol. 9, no. Nov, pp. 2579–2605, 2008
2008
Earlier work this paper cites.
M. F. Land, “Vision, eye movements, and natural behavior,” Visual Neuroscience , vol. 26, no. 1, pp. 51–62, 2009
2009
Earlier work this paper cites.
T. Dodson, N. Mattei, and J. Goldsmith, “A natural language argumentation interface for explanation generation in markov decision processes,” in International Conference on Algorithmic Decision Theory . Springer, 2011, pp. 42–55
2011
Earlier work this paper cites.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling, “The arcade learning environment: An evaluation platform for general agents,” Journal of Artificial Intelligence Research , vol. 47, pp. 253–279, 2013
2013
Earlier work this paper cites.
K. Simonyan, A. Vedaldi, and A. Zisserman, “Deep inside convolutional networks: Visualising image classification models and saliency maps,” in Workshop Proceedings of the International Conference on Learning Representations , 2014
2014
Earlier work this paper cites.
M. D. Zeiler and R. Fergus, “Visualizing and understanding convolutional networks,” in Proceedings of the European Conference on Computer Vision . Springer, 2014, pp. 818–833
2014
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, p. 529, 2015
2015
Earlier work this paper cites.
U. D. Gupta, E. Talvitie, and M. Bowling, “Policy tree: Adaptive representation for policy gradient,” in Twenty-Ninth AAAI Conference on Artificial Intelligence , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: Convolutional networks for biomedical image segmentation,” in International Conference on Medical Image Computing and Computer-Assisted Intervention . Springer, 2015, pp. 234–241
2015
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of go with deep neural networks and tree search,” Nature , vol. 529, no. 7587, p. 484, 2016
2016
Earlier work this paper cites.
M. T. Ribeiro, S. Singh, and C. Guestrin, “”why should i trust you?” explaining the predictions of any classifier,” in Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining , 2016, pp. 1135–1144
2016
Earlier work this paper cites.
A. Binder, G. Montavon, S. Lapuschkin, K.-R. Müller, and W. Samek, “Layer-wise relevance propagation for neural networks with local renormalization layers,” in International Conference on Artificial Neural Networks . Springer, 2016, pp. 63–71
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
H. Van Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning,” in Thirtieth AAAI Conference on Artificial Intelligence , 2016
2016
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” in International Conference on Learning Representations , 2016
2016
Cited alongside, same era.
N. Mayer, E. Ilg, P. Hausser, P. Fischer, D. Cremers, A. Dosovitskiy, and T. Brox, “A large dataset to train convolutional networks for disparity, optical flow, and scene flow estimation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 4040–4048
2016
Cited alongside, same era.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Cited alongside, same era.
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. J. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu et al. , “Learning to navigate in complex environments,” in Proceedings of the International Conference on Learning Representations , 2017
2017
J. Waa, J. v. Diggelen, K. Bosch, and M. Neerincx, “Contrastive explanations for reinforcement learning in terms of expected consequences,” in Proceedings of the Workshop on Explainable AI on the IJCAI , 2018
2018
Later among the works it cites.
J. Zhang, S. A. Bargal, Z. Lin, J. Brandt, X. Shen, and S. Sclaroff, “Top-down neural attention by excitation backprop,” International Journal of Computer Vision , vol. 126, no. 10, pp. 1084–1102, 2018
2018
Later among the works it cites.
V. Petsiuk, A. Das, and K. Saenko, “Rise: Randomized input sampling for explanation of black-box models,” in British Machine Vision Conference , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra, “Grad-cam: Visual explanations from deep networks via gradient-based localization,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 618–626
2017
Cited alongside, same era.
S. M. Lundberg and S.-I. Lee, “A unified approach to interpreting model predictions,” in Advances in Neural Information Processing Systems , 2017, pp. 4765–4774
2017
Cited alongside, same era.
D. Bau, B. Zhou, A. Khosla, A. Oliva, and A. Torralba, “Network dissection: Quantifying interpretability of deep visual representations,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 6541–6549
2017
Cited alongside, same era.
B. Hayes and J. A. Shah, “Improving robot controller transparency through autonomous policy explanation,” in 12th ACM/IEEE International Conference on Human-Robot Interaction , 2017, pp. 303–312
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
R. C. Fong and A. Vedaldi, “Interpretable explanations of black boxes by meaningful perturbation,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 3429–3437
2017
Cited alongside, same era.
P. Dabkowski and Y. Gal, “Real time image saliency for black box classifiers,” in Advances in Neural Information Processing Systems , 2017, pp. 6967–6976
2017
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International Conference on Machine Learning , 2018
2018
Later among the works it cites.
S. Fujimoto, H. van Hoof, and D. Meger, “Addressing function approximation error in actor-critic methods,” in International Conference on Machine Learning , 2018
2018
Later among the works it cites.
W. Shi, S. Song, H. Wu, Y.-C. Hsu, C. Wu, and G. Huang, “Regularized anderson acceleration for off-policy deep reinforcement learning,” in Advances in Neural Information Processing Systems , 2019, pp. 10 231–10 241
2019
Later among the works it cites.
Q. Cao, X. Liang, B. Li, and L. Lin, “Interpretable visual question answering by reasoning on dependency trees,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2019
2019
Later among the works it cites.
M. Monfort, A. Andonian, B. Zhou, K. Ramakrishnan, S. A. Bargal, T. Yan, L. Brown, Q. Fan, D. Gutfreund, C. Vondrick et al. , “Moments in time dataset: one million videos for event understanding,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 42, no. 2, pp. 502–508, 2019
2019
Later among the works it cites.
H. Liu, R. Wang, S. Shan, and X. Chen, “What is tabby? interpretable model decisions by learning attribute-based classification criteria,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2019
2019
Later among the works it cites.
R. Guidotti, A. Monreale, S. Ruggieri, F. Turini, F. Giannotti, and D. Pedreschi, “A survey of methods for explaining black box models,” ACM Computing Surveys (CSUR) , vol. 51, no. 5, p. 93, 2019
2019
Later among the works it cites.
R. M. Annasamy and K. Sycara, “Towards better interpretability in deep q-networks,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, 2019, pp. 4561–4569
2019
Later among the works it cites.
2019
Later among the works it cites.
A. Mott, D. Zoran, M. Chrzanowski, D. Wierstra, and D. Jimenez Rezende, “Towards interpretable reinforcement learning using attention augmented agents,” in Advances in Neural Information Processing Systems , 2019, pp. 12 350–12 359
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Ancona, C. Öztireli, and M. Gross, “Explaining deep neural networks with a polynomial time algorithm for shapley values approximation,” in International Conference on Machine Learning , 2019
2019
Later among the works it cites.
R. Fong, M. Patrick, and A. Vedaldi, “Understanding deep networks via extremal perturbations and smooth masks,” in Proceedings of the IEEE International Conference on Computer Vision , 2019, pp. 2950–2958
2019
Later among the works it cites.
2019
Later among the works it cites.
W. Shi, S. Song, and C. Wu, “Soft policy gradient method for maximum entropy deep reinforcement learning,” in Proceedings of the 28th International Joint Conference on Artificial Intelligence , 2019, pp. 3425–3431
2019
Later among the works it cites.
W. Wang, J. Shen, X. Lu, S. C. Hoi, and H. Ling, “Paying attention to video object pattern understanding,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2020
2020
Closest in time.