Decoupling feature extraction from policy learning: assessing benefits of state representation learning in goal based robotics (2019)
Raffin, A. et al · 1901
Earlier work this paper cites.
Human-centered tools for coping with imperfect algorithms during medical decision-making
Original
Cai, C. J. et al · 1902
Earlier work this paper cites.
Explainable reinforcement learning through a causal lens (2019)
Madumal, P., Miller, T., Sonenberg, L. & Vetere, F · 1905
Earlier work this paper cites.
Discorl: Continual reinforcement learning via policy distillation (2019)
Traoré, R. et al · 1907
Earlier work this paper cites.
Sentimate: Learning to play chess through natural language processing (2019)
Kamlish, I., Chocron, I. B. & McCarthy, N · 1907
Earlier work this paper cites.
Efficient saliency maps for explainable ai (2020)
Mundhenk, T. N., Chen, B. Y. & Friedland, G · 1911
Earlier work this paper cites.
Explain your move: Understanding agent actions using specific and relevant feature attribution (2020)
Puri, N. et al · 1912
Earlier work this paper cites.
Positive matrix factorization: A non-negative factor model with optimal utilization of error estimates of data values
Paatero, P. & Tapper, U · 1994
Earlier work this paper cites.
Case-based evaluation in computer chess
Kerner, Y · 1995
Earlier work this paper cites.
Learning strategies for explanation patterns: Basic game patterns with application to chess
HaCohen-Kerner, Y · 1995
Earlier work this paper cites.
Learning the parts of objects by non-negative matrix factorization
Lee, D. D. & Seung, H. S · 1999
Earlier work this paper cites.
Dream architecture: a developmental approach to open-ended learning in robotics (2020)
Doncieux, S. et al · 2005
Earlier work this paper cites.
Gradient alignment in deep neural networks
Original
Srinivas, S. & Fleuret, F · 2006
Earlier work this paper cites.
Explainability in deep reinforcement learning (2020)
Heuillet, A., Couthouis, F. & Díaz-Rodríguez, N · 2008
Earlier work this paper cites.
Learning personalized models of human behavior in chess (2021)
McIlroy-Young, R., Wang, R., Sen, S., Kleinberg, J. & Anderson, A · 2008
Earlier work this paper cites.
Visualizing data using t-sne
Van der Maaten, L. & Hinton, G · 2008
Earlier work this paper cites.
Assessing game balance with alphazero: Exploring alternative rule sets in chess (2020)
Tomašev, N., Paquet, U., Hassabis, D. & Kramnik, V · 2009
Earlier work this paper cites.
Domain-level explainability – a challenge for creating trust in superhuman ai strategies (2020)
Andrulis, J., Meyer, O., Schott, G., Weinbach, S. & Gruhn, V · 2011
Earlier work this paper cites.
Debugging tests for model explanations
Original
Adebayo, J., Muelly, M., Liccardi, I. & Kim, B · 2011
Earlier work this paper cites.
Hierarchical sub-task decomposition for reinforcement learning of multi-robot delivery mission
Kawano, H · 2013
Earlier work this paper cites.
Chess q & a : Question answering on chess games (2015)
Cirik, V., Morency, L.-P. & Hovy, E · 2015
Earlier work this paper cites.
Understanding intermediate layers using linear classifier probes
Original
Alain, G. & Bengio, Y · 2016
Earlier work this paper cites.
Towards deep symbolic reinforcement learning (2016)
Garnelo, M., Arulkumaran, K. & Shanahan, M · 2016
Earlier work this paper cites.
Graying the black box: Understanding DQNs
Zahavy, T., Ben-Zrihem, N. & Mannor, S · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S. & Sun, J · 2016
Earlier work this paper cites.
Network dissection: Quantifying interpretability of deep visual representations
Bau, D., Zhou, B., Khosla, A., Oliva, A. & Torralba, A · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Original
Doshi-Velez, B., Finale; Kim · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
Lundberg, S. M. & Lee, S.-I · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks
Sundararajan, M., Taly, A. & Yan, Q · 2017
Earlier work this paper cites.
Interpretable explanations of black boxes by meaningful perturbation
Original
Fong, R. & Vedaldi, A · 2017
Earlier work this paper cites.
Grad-cam: Visual explanations from deep networks via gradient-based localization
Selvaraju, R. R. et al · 2017
Earlier work this paper cites.