Fetching the paper…
Reading the bibliography…
Reinforcement learning (RL) systems can be complex and non-interpretable, making it challenging for non-AI experts to understand or intervene in their decisions.
Chan, W.T., Chow, Y.K., Liu, L.F.: Neural network: An alternative to pile driving formulas. Computers and Geotechnics 17
1995
Earlier work this paper cites.
Ng, A.Y., Harada, D., Russell, S.: Policy invariance under reward transformations: Theory and application to reward shaping. In: Icml, vol. 99, pp. 278–287 (1999). Citeseer
1999
Earlier work this paper cites.
Braun, V., Clarke, V.: Using thematic analysis in psychology. Qualitative research in psychology 3
2006
Earlier work this paper cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., et al
2015
Earlier work this paper cites.
Mnih, V., Badia, A.P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., Kavukcuoglu, K.: Asynchronous methods for deep reinforcement learning. In: International Conference on Machine Learning, pp. 1928–1937 (2016). PMLR
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
Sutton, R.S., Barto, A.G.: Reinforcement Learning: An Introduction. A Bradford Book, Cambridge, MA, USA (2018)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Chevalier-Boisvert, M., Willems, L., Pal, S.: Minimalistic Gridworld Environment for OpenAI Gym. GitHub (2018). https://github.com/maximecb/gym-minigrid
2018
Earlier work this paper cites.
Miller, T.: Explanation in artificial intelligence: Insights from the social sciences. Artificial intelligence 267
2019
Earlier work this paper cites.
Anjomshoae, S., Najjar, A., Calvaresi, D., Främling, K.: Explainable Agents and Robots: Results from a Systematic Literature Review. In: Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems. AAMAS ’19, pp. 1078–1088. International Foundation for Autonomous Agents and Multiagent Systems, Richland, SC (2019)
2019
Earlier work this paper cites.
Juozapaitis, Z., Koul, A., Fern, A., Erwig, M., Doshi-Velez, F.: Explainable reinforcement learning via reward decomposition. In: IJCAI/ECAI Workshop on Explainable Artificial Intelligence (2019)
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Ehsan, U., Tambwekar, P., Chan, L., Harrison, B., Riedl, M.O.: Automated rationale generation: a technique for explainable ai and its effects on human perceptions. In: Proceedings of the 24th International Conference on Intelligent User Interfaces, pp. 263–274 (2019)
2019
Earlier work this paper cites.
Alharin, A., Doan, T.-N., Sartipi, M.: Reinforcement learning interpretation methods: A survey. IEEE Access 8
2020
Cited alongside, same era.
Puiutta, E., Veith, E.M.S.P.: Explainable Reinforcement Learning: A Survey. In: Holzinger, A., Kieseberg, P., Tjoa, A.M., Weippl, E. (eds.) Machine Learning and Knowledge Extraction. Lecture Notes in Computer Science, pp. 77–95. Springer, Cham (2020). https://doi.org/10.1007/978-3-030-57321-8_5
2020
Cited alongside, same era.
Das, D., Chernova, S.: Leveraging rationales to improve human task performance. In: Proceedings of the 25th International Conference on Intelligent User Interfaces, pp. 510–518 (2020)
2020
Cited alongside, same era.
Madumal, P., Miller, T., Sonenberg, L., Vetere, F.: Explainable reinforcement learning through a causal lens. In: Proceedings of the AAAI Conference on Artificial Intelligence, vol. 34, pp. 2493–2500 (2020)
2020
Cited alongside, same era.
Wells, L., Bednarz, T.: Explainable AI and Reinforcement Learning—A Systematic Review of Current Approaches and Trends. Frontiers in Artificial Intelligence 4
2022
Closest in time.
Heuillet, A., Couthouis, F., Díaz-Rodríguez, N.: Explainability in deep reinforcement learning. Knowledge-Based Systems 214
2022
Closest in time.
Chakraborti, T., Sreedharan, S., Kambhampati, S.: The Emerging Landscape of Explainable Automated Planning & Decision Making. In: Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence, pp. 4803–4811. International Joint Conferences on Artificial Intelligence Organization, Yokohama, Japan (2020). https://doi.org/10.24963/ijcai.2020/669 . https://www.ijcai.org/proceedings/2020/669
2022
Closest in time.
Sreedharan, S., Soni, U., Verma, M., Srivastava, S., Kambhampati, S.: Bridging the gap: Providing post-hoc symbolic explanations for sequential decision-making problems with inscrutable representations. In: International Conference on Learning Representations (2022). https://openreview.net/forum?id=o-1v9hdSult
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
Miles, M.B., Huberman, A.M., Saldaña, J.: Qualitative Data Analysis : a Methods Sourcebook, Fourth edition edn. SAGE Los Angeles, Los Angeles (2020)
2020
Cited alongside, same era.
Zelvelder, A.E., Westberg, M., Främling, K.: Assessing explainability in reinforcement learning. In: International Workshop on Explainable, Transparent Autonomous Agents and Multi-Agent Systems, pp. 223–240 (2021). Springer
2021
Cited alongside, same era.
Chakraborti, T., Sreedharan, S., Kambhampati, S.: The emerging landscape of explainable automated planning & decision making. In: Proceedings of the Twenty-Ninth International Conference on International Joint Conferences on Artificial Intelligence, pp. 4803–4811 (2021)
2021
Cited alongside, same era.
Olson, M.L., Khanna, R., Neal, L., Li, F., Wong, W.-K.: Counterfactual state explanations for reinforcement learning agents via generative deep learning. Artificial Intelligence 295
2021
Cited alongside, same era.
Dazeley, R., Vamplew, P., Foale, C., Young, C., Aryal, S., Cruz, F.: Levels of explainable artificial intelligence for human-aligned conversational explanations. Artificial Intelligence 299
2021
Cited alongside, same era.
Hafner, D.: Benchmarking the spectrum of agent capabilities. arXiv preprint arXiv:2109.06780 (2021)
2021
Cited alongside, same era.
Zhou, J., Gandomi, A.H., Chen, F., Holzinger, A.: Evaluating the quality of machine learning explanations: A survey on methods and metrics. Electronics 10
2021
Cited alongside, same era.
2022
Closest in time.
Frost, J., Watkins, O., Weiner, E., Abbeel, P., Darrell, T., Plummer, B., Saenko, K.: Explaining reinforcement learning policies through counterfactual trajectories. ICML Workshop on Human in the Loop Learning (HILL) (2022)
2022
Closest in time.
Willems, L.: RL Starter Files for MiniGrid (2018). https://github.com/lcswillems/rl-starter-files/tree/4205e05b7905fec16519bc0802596673d86af018 Accessed 2022-09-28
2022
Closest in time.
Qian, P., Unhelkar, V.: Evaluating the role of interactivity on improving transparency in autonomous agents. In: Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems, pp. 1083–1091 (2022)
2022
Closest in time.
Stanić, A., Tang, Y., Ha, D., Schmidhuber, J.: Learning to Generalize with Object-centric Agents in the Open World Survival Game Crafter (2022)
2022
Closest in time.
Milani, S., Topin, N., Veloso, M., Fang, F.: Explainable reinforcement learning: A survey and comparative review. ACM Computing Surveys (2023) https://doi.org/10.1145/3616864
2023
Closest in time.
Septon, Y., Huber, T., André, E., Amir, O.: Integrating policy summaries with reward decomposition for explaining reinforcement learning agents. In: International Conference on Practical Applications of Agents and Multi-Agent Systems, pp. 320–332 (2023). Springer
2023
Closest in time.
Huber, T., Demmler, M., Mertes, S., Olson, M.L., André, E.: Ganterfactual-rl: Understanding reinforcement learning agents’ strategies through visual counterfactual explanations. In: Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems, pp. 1097–1106 (2023)
2023
Closest in time.
Hoffman, R.R., Mueller, S.T., Klein, G., Litman, J.: Measures for explainable ai: Explanation goodness, user satisfaction, mental models, curiosity, trust, and human-ai performance. Frontiers in Computer Science 5
2023
Closest in time.
Das, D., Chernova, S., Kim, B.: State2explanation: concept-based explanations to benefit agent learning and user understanding. In: Proceedings of the 37th International Conference on Neural Information Processing Systems. NIPS ’23. Curran Associates Inc., Red Hook, NY, USA (2024)
2024
Closest in time.
Tambwekar, P., Gombolay, M.: Towards reconciling usability and usefulness of policy explanations for sequential decision-making systems. Frontiers in Robotics and AI 11
2024
Closest in time.