Fetching the paper…
Reading the bibliography…
Broad Explainable Artificial Intelligence moves away from interpreting individual decisions based on a single datum and aims to provide integrated explanations from multiple machine learning algorithms into a coherent explanation of an agent's behaviour that is aligned to the communication needs of the explainee.
Complementary reinforcement learning towards explainable agents
Lee, J.H., 2019 · 1901
Earlier work this paper cites.
Counterfactual visual explanations
Goyal, Y., Wu, Z., Ernst, J., Batra, D., Parikh, D., Lee, S., 2019 · 1904
Earlier work this paper cites.
Exploring computational user models for agent policy summarization
Lage, I., Lifschitz, D., Doshi-Velez, F., Amir, O., 2019a · 1905
Earlier work this paper cites.
Explainable reinforcement learning through a causal lens
Madumal, P., Miller, T., Sonenberg, L., Vetere, F., 2019 · 1905
Earlier work this paper cites.
Generating counterfactual and contrastive explanations using SHAP
Rathi, S., 2019 · 1906
Earlier work this paper cites.
Uncertainty-aware model-based policy optimization
Vuong, T.L., Tran, K., 2019 · 1906
Earlier work this paper cites.
Towards explainable AI planning as a service
Cashmore, M., Collins, A., Krarup, B., Krivic, S., Magazzeni, D., Smith, D., 2019 · 1908
Earlier work this paper cites.
Graying the black box: Understanding dqns, in: International Conference on Machine Learning, pp. 1899–1908
Zahavy, T., Ben-Zrihem, N., Mannor, S., 2016 · 1908
Earlier work this paper cites.
Generating justifications for norm-related agent decisions
Kasenberg, D., Roque, A., Thielstrom, R., Chita-Tegmark, M., Scheutz, M., 2019a · 1911
Earlier work this paper cites.
Engaging in dialogue about an agent’s norms and behaviors
Kasenberg, D., Roque, A., Thielstrom, R., Scheutz, M., 2019b · 1911
Earlier work this paper cites.
Exploratory not explanatory: Counterfactual analysis of saliency maps for deep rl
Atrey, A., Clary, K., Jensen, D., 2019 · 1912
Earlier work this paper cites.
The psychology of interpersonal relations
Heider, F., 1958 · 1958
Earlier work this paper cites.
From acts to dispositions the attribution process in person perception, in: Advances in experimental social psychology. Elsevier. volume 2, pp. 219–266
Jones, E.E., Davis, K.E., 1965 · 1965
Earlier work this paper cites.
Attribution theory in social psychology., in: Nebraska symposium on motivation, University of Nebraska Press
Kelley, H.H., 1967 · 1967
Earlier work this paper cites.
The processes of causal attribution
Kelley, H.H., 1973 · 1973
Earlier work this paper cites.
A model of inexact reasoning in medicine
Shortliffe, E.H., Buchanan, B.G., 1975 · 1975
Earlier work this paper cites.
The Question Of Animal Awareness: Evolutionary Continuity Of Mental Experience
Griffin, D.R., 1976 · 1976
Earlier work this paper cites.
Production rules as a representation for a knowledge-based consultation program
Davis, R., Buchanan, B., Shortliffe, E., 1977 · 1977
Earlier work this paper cites.
Explaining emotions
Rorty, A.O., 1978 · 1978
Earlier work this paper cites.
XPLAIN: A system for creating and explaining expert consulting programs
Swartout, W.R., 1983 · 1983
Earlier work this paper cites.
Pretense and representation: The origins of "theory of mind."
Leslie, A.M., 1987 · 1987
Earlier work this paper cites.
Explanation: the role of control strategies and deep models
Chandrasekaran, B., Tanner, M.C., Josephson, J.R., 1988 · 1988
Earlier work this paper cites.
Utility-directed presentation of simulation results, in: Proceedings of the Annual Symposium on Computer Application in Medical Care, American Medical Informatics Association. p. 292
McLaughlin, J., 1988 · 1988
Earlier work this paper cites.
Explanatory coherence
Thagard, P., 1989 · 1989
Earlier work this paper cites.
How Monkeys See The World: Inside the mind of another species
Cheney, D.L., Seyfarth, R.M., 1990 · 1990
Earlier work this paper cites.
Contrastive explanation
Lipton, P., 1990 · 1990
Earlier work this paper cites.
Social cognition
Fiske, S.T., Taylor, S.E., 1991 · 1991
Earlier work this paper cites.
Incomplete information and deception in multi-agent negotiation., in: IJCAI, pp. 225–231
Zlotkin, G., Rosenschein, J.S., 1991 · 1991
Earlier work this paper cites.
Learning to achieve goals, in: In Proceedings of the Thirteenth International Joint Conference on Artificial Intelligence, Morgan Kaufmann. pp. 1094–1098
Kaelbling, L.P., 1993 · 1993
Earlier work this paper cites.
Explaining emotions
O’Rorke, P., Ortony, A., 1994 · 1994
Earlier work this paper cites.
Survey and critique of techniques for extracting rules from trained artificial neural networks
Andrews, R., Diederich, J., Tickle, A.B., 1995 · 1995
Earlier work this paper cites.
BDI agents: from theory to practice., in: ICMAS, pp. 312–319
Rao, A.S., Georgeff, M.P., et al., 1995 · 1995
Earlier work this paper cites.
Reinforcement learning with soft state aggregation, in: Advances in neural information processing systems, pp. 361–368
Singh, S.P., Jaakkola, T., Jordan, M.I., 1995 · 1995
Earlier work this paper cites.
Explanation in probabilistic systems: Is it feasible? will it work, Citeseer
Druzdzel, M.J., 1996 · 1996
Earlier work this paper cites.
A density-based algorithm for discovering clusters in large spatial databases with noise
Ester, M., Kriegel, H.P., Sander, J., Xu, X., et al., 1996 · 1996
Earlier work this paper cites.
Controlling ourselves, controlling our world: Psychology’s role in understanding positive and negative consequences of seeking and gaining control
Shapiro Jr, D.H., Schwartz, C.E., Astin, J.A., 1996 · 1996
Earlier work this paper cites.
An efficient algorithm for temporal abduction, in: Congress of the Italian Association for Artificial Intelligence, Springer. pp. 195–206
Brusoni, V., Console, L., Terenziani, P., Dupré, D.T., 1997 · 1997
Earlier work this paper cites.
Multitask learning
Caruana, R., 1997 · 1997
Earlier work this paper cites.
Modelling social action for AI agents
Castelfranchi, C., 1998 · 1998
Earlier work this paper cites.
A model of emotion-driven choice
Elliott, R., 1998 · 1998
Earlier work this paper cites.
Decision making in qualitative influence diagrams., in: FLAIRS Conference, pp. 410–414
Renooij, S., Van Der Gaag, L.C., 1998 · 1998
Earlier work this paper cites.
How people explain behavior: A new theoretical framework
Malle, B.F., 1999 · 1999
Earlier work this paper cites.
Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning
Sutton, R.S., Precup, D., Singh, S., 1999 · 1999
Earlier work this paper cites.
Graphical explanation in bayesian networks, in: International Symposium on Medical Data Analysis, Springer. pp. 122–129
Lacave, C., Atienza, R., Díez, F.J., 2000 · 2000
Earlier work this paper cites.
Conceptual structure and social functions of behavior explanations: Beyond person–situation attributions
Malle, B.F., Knobe, J., O’Laughlin, M.J., Pearce, G.E., Nelson, S.E., 2000 · 2000
Earlier work this paper cites.
Robot learning driven by emotions
Gadanho, S.C., Hallam, J., 2001 · 2001
Earlier work this paper cites.
Hierarchical reinforcement learning as a model of human task interleaving
Gebhardt, C., Oulasvirta, A., Hilliges, O., 2020 · 2001
Earlier work this paper cites.
A review on generative adversarial networks: Algorithms, theory, and applications
Gui, J., Sun, Z., Wen, Y., Tao, D., Ye, J., 2020 · 2001
Earlier work this paper cites.
Reinforcement learning in dynamic environments using instantiated information, in: Machine Learning: Proceedings of the Eighteenth International Conference (ICML2001), pp. 585–592
Wiering, M.A., 2001 · 2001
Earlier work this paper cites.
The emerging landscape of explainable AI planning and decision making
Chakraborti, T., Sreedharan, S., Kambhampati, S., 2020 · 2002
Earlier work this paper cites.
Multiple model-based reinforcement learning
Doya, K., Samejima, K., Katagiri, K.i., Kawato, M., 2002 · 2002
Earlier work this paper cites.
A review of explanation methods for bayesian networks
Lacave, C., Díez, F.J., 2002 · 2002
Earlier work this paper cites.
Temporal-adaptive hierarchical reinforcement learning
Zhou, W.J., Yu, Y., 2020 · 2002
Earlier work this paper cites.
Optimal decision explanation by extracting regularity patterns, in: Coenen F., Preece A., Macintosh A. (eds) Research and Development in Intelligent Systems XX. SGAI 2003, Springer. pp. 283–294
Bielza, C., Fernández del Pozo, J.A., Lucas, P., 2003 · 2003
Earlier work this paper cites.
Towards transparent robotic planning via contrastive explanations
Chen, S., Boggess, K., Feng, L., 2020 · 2003
Earlier work this paper cites.
Self-supervised discovering of causal features: Towards interpretable reinforcement learning
Shi, W., Wang, Z., Song, S., Huang, G., 2020 · 2003
Earlier work this paper cites.
Intrinsically motivated learning of hierarchical collections of skills, in: Proceedings of the 3rd International Conference on Development and Learning, pp. 112–19
Barto, A.G., Singh, S., Chentanez, N., 2004 · 2004
Earlier work this paper cites.
Explainable goal-driven agents and robots–a comprehensive review and new framework
Sado, F., Loo, C.K., Kerzel, M., Wermter, S., 2020 · 2004
Earlier work this paper cites.
Tradeoff-focused contrastive explanation for MDP planning
Sukkerd, R., Simmons, R., Garlan, D., 2020 · 2004
Earlier work this paper cites.
Intrinsically motivated reinforcement learning, in: Advances in neural information processing systems, pp. 1281–1288
Chentanez, N., Barto, A.G., Singh, S.P., 2005 · 2005
Earlier work this paper cites.
Incorporating if… then… personality signatures in person perception: beyond the person-situation dichotomy
Kammrath, L.K., Mendoza-Denton, R., Mischel, W., 2005 · 2005
Earlier work this paper cites.
Robust reinforcement learning
Morimoto, J., Doya, K., 2005 · 2005
Earlier work this paper cites.
Explainable reinforcement learning: A survey
Puiutta, E., Veith, E., 2020 · 2005
Earlier work this paper cites.
A survey of reinforcement learning in relational domains
Van Otterlo, M., 2005 · 2005
Earlier work this paper cites.
Explanations and recommendations for temporal inconsistencies
Bresina, J.L., Morris, P.H., 2006 · 2006
Earlier work this paper cites.
Explainable robotic systems: Understanding goal-driven actions in a reinforcement learning scenario
Cruz, F., Dazeley, R., Vamplew, P., 2020 · 2006
Earlier work this paper cites.
How the mind explains behavior: Folk explanations, meaning, and social interaction
Malle, B.F., 2006 · 2006
Earlier work this paper cites.
Meta-explanation in a constraint satisfaction solver, in: Information Processing and Management of Uncertainty in Knowledge-based Systems IPMU, Citeseer. pp. 1118–1125
Pitrat, J., et al., 2006 · 2006
Earlier work this paper cites.
PersonisAD: Distributed, active, scrutable model framework for context-aware services, in: International Conference on Pervasive Computing, Springer. pp. 55–72
Assad, M., Carmichael, D.J., Kay, J., Kummerfeld, B., 2007 · 2007
Earlier work this paper cites.
An MDP approach for explanation generation
Elizalde, F., Sucar, L.E., Reyes, A., Debuen, P., 2007 · 2007
Earlier work this paper cites.
Simplicity and probability in causal explanation
Lombrozo, T., 2007 · 2007
Earlier work this paper cites.
The effects of transparency on trust in and acceptance of a content-based art recommender
Cramer, H., Evers, V., Ramlal, S., Van Someren, M., Rutledge, L., Stash, N., Aroyo, L., Wielinga, B., 2008 · 2008
Earlier work this paper cites.
Epistemological approach to the process of practice
Dazeley, R., Kang, B.H., 2008 · 2008
Earlier work this paper cites.
Policy explanation in factored markov decision processes
Elizalde, F., 2008 · 2008
Earlier work this paper cites.
Explainability in deep reinforcement learning
Heuillet, A., Couthouis, F., Rodríguez, N.D., 2020 · 2008
Earlier work this paper cites.
Emotion-driven reinforcement learning, in: Proceedings of the Annual Meeting of the Cognitive Science Society
Marinier, R.P., Laird, J.E., 2008 · 2008
Earlier work this paper cites.
Hierarchically organized behavior and its neural foundations: A reinforcement learning perspective
Botvinick, M.M., Niv, Y., Barto, A.G., 2009 · 2009
Earlier work this paper cites.
Generating explanations based on markov decision processes, in: Mexican International Conference on Artificial Intelligence, Springer. pp. 51–62
Elizalde, F., Sucar, E., Noguez, J., Reyes, A., 2009 · 2009
Earlier work this paper cites.
Minimal sufficient explanations for factored markov decision processes, in: Nineteenth International Conference on Automated Planning and Scheduling
Khan, O.Z., Poupart, P., Black, J.P., 2009 · 2009
Earlier work this paper cites.
A computational unification of cognitive behavior and emotion
Marinier III, R.P., Laird, J.E., Lewis, R.L., 2009 · 2009
Earlier work this paper cites.
Reinforcement learning and dynamic programming using function approximators. volume 39
Busoniu, L., Babuska, R., De Schutter, B., Ernst, D., 2010 · 2010
Earlier work this paper cites.
Explanation versus meta-explanation: What makes a case more convincing, in: FLAIRS Conference
Galitsky, B.A., de la Rosa i Esteva, J.L., Kovalerchuk, B., 2010 · 2010
Earlier work this paper cites.
Cognitive modelling of human social signals., in: SSPW@ MM, pp. 21–26
Poggi, I., D’Errico, F., 2010 · 2010
Cited alongside, same era.
The many faces of deception
Sakama, C., Caminada, M., 2010 · 2010
Cited alongside, same era.
Explaining conditions for reinforcement learning behaviors from real and imagined data
Acharya, A., Russell, R., Ahmed, N.R., 2020 · 2011
Cited alongside, same era.
Domain-level explainability–a challenge for creating trust in superhuman ai strategies
Andrulis, J., Meyer, O., Schott, G., Weinbach, S., Gruhn, V., 2020 · 2011
Cited alongside, same era.
Behavioral game theory: Experiments in strategic interaction
Camerer, C.F., 2011 · 2011
Cited alongside, same era.
Robot transparency, trust and utility
Wortham, R.H., Theodorou, A., 2017 · 2017
Later among the works it cites.
Trends and trajectories for explainable, accountable and intelligible systems: An HCI research agenda, in: Proceedings of the SIGCHI Conference on Human Factors in Computing Systems. CHI’18
Abdul, A., Vermeulen, J., Wang, D., Lim, B.Y., Kankanhalli, M., 2018 · 2018
Later among the works it cites.
Peeking inside the black-box: A survey on explainable artificial intelligence (XAI)
Adadi, A., Berrada, M., 2018 · 2018
Later among the works it cites.
Highlights: Summarizing agent behavior to people, in: Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems, International Foundation for Autonomous Agents and Multiagent Systems. pp. 1168–1176
Amir, D., Amir, O., 2018 · 2018
Later among the works it cites.
Agent strategy summarization, in: Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems, International Foundation for Autonomous Agents and Multiagent Systems. pp. 1203–1207
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A natural language argumentation interface for explanation generation in markov decision processes, in: International Conference on Algorithmic DecisionTheory, Springer. pp. 42–55
Dodson, T., Mattei, N., Goldsmith, J., 2011 · 2011
Cited alongside, same era.
Constructing and revising commonsense science explanations: A metareasoning approach, in: 2011 AAAI Fall Symposium Series
Friedman, S., Forbus, K.D., Sherin, B., 2011 · 2011
Cited alongside, same era.
The current state of normative agent-based systems
Hollander, C.D., Wu, A.S., 2011 · 2011
Cited alongside, same era.
Asp-prolog for negotiation among dishonest agents, in: International Conference on Logic Programming and Nonmonotonic Reasoning, Springer. pp. 331–344
Nguyen, N.H., Son, T.C., Pontelli, E., Sakama, C., 2011 · 2011
Cited alongside, same era.
A logical formulation for negotiation among dishonest agents, in: Twenty-Second International Joint Conference on Artificial Intelligence
Sakama, C., Tran, S.C., Pontelli, E., 2011 · 2011
Cited alongside, same era.
Programming mental state abduction, in: The 10th International Conference on Autonomous Agents and Multiagent Systems-Volume 1, International Foundation for Autonomous Agents and Multiagent Systems. pp. 301–308
Sindlar, M., Dastani, M., Meyer, J.J., 2011 · 2011
Cited alongside, same era.
Empirical evaluation methods for multiobjective reinforcement learning algorithms
Vamplew, P., Dazeley, R., Berry, A., Issabekov, R., Dekker, E., 2011 · 2011
Cited alongside, same era.
Amir, O., Doshi-Velez, F., Sarne, D., 2018 · 2018
Later among the works it cites.
Explanations for temporal recommendations
Bharadhwaj, H., Joshi, S., 2018 · 2018
Later among the works it cites.
Explaining unsolvable planning tasks
Bongartz, I.N., 2018 · 2018
Later among the works it cites.
Explaining image classifiers by counterfactual generation
Chang, C.H., Creager, E., Goldenberg, A., Duvenaud, D., 2018 · 2018
Later among the works it cites.
Model-based reinforcement learning via meta-policy optimization
Clavera, I., Rothfuss, J., Schulman, J., Fujita, Y., Asfour, T., Abbeel, P., 2018 · 2018
Later among the works it cites.
Explanations based on the missing: Towards contrastive explanations with pertinent negatives, in: Advances in neural information processing systems, pp. 592–603
Dhurandhar, A., Chen, P.Y., Luss, R., Tu, C.C., Ting, P., Shanmugam, K., Das, P., 2018 · 2018
Later among the works it cites.
Explaining deep adaptive programs via reward decomposition, in: IJCAI/ECAI Workshop on Explainable Artificial Intelligence
Erwig, M., Fern, A., Murali, M., Koul, A., 2018 · 2018
Later among the works it cites.
Explaining explanations: An overview of interpretability of machine learning, in: 2018 IEEE 5th International Conference on data science and advanced analytics (DSAA), IEEE. pp. 80–89
Gilpin, L.H., Bau, D., Yuan, B.Z., Bajwa, A., Specter, M., Kagal, L., 2018 · 2018
Later among the works it cites.
Temporal difference variational auto-encoder
Gregor, K., Papamakarios, G., Besse, F., Buesing, L., Weber, T., 2018 · 2018
Later among the works it cites.
Understanding individual decisions of cnns via contrastive backpropagation, in: Asian Conference on Computer Vision, Springer. pp. 119–134
Gu, J., Yang, Y., Tresp, V., 2018 · 2018
Later among the works it cites.
Social gan: Socially acceptable trajectories with generative adversarial networks, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 2255–2264
Gupta, A., Johnson, J., Fei-Fei, L., Savarese, S., Alahi, A., 2018 · 2018
Later among the works it cites.
Interpretable policies for reinforcement learning by genetic programming
Hein, D., Udluft, S., Runkler, T.A., 2018 · 2018
Later among the works it cites.
Understandable robots-what, why, and how
Hellström, T., Bensch, S., 2018 · 2018
Later among the works it cites.
Establishing appropriate trust via critical states, in: 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), IEEE. pp. 3929–3936
Huang, S.H., Bhatia, K., Abbeel, P., Dragan, A.D., 2018 · 2018
Later among the works it cites.
The role of trust in human-robot interaction, in: Foundations of trusted autonomy. Springer, Cham, pp. 135–159
Lewis, M., Sycara, K., Walker, P., 2018 · 2018
Later among the works it cites.
Contrastive explanation: A structural-model approach
Miller, T., 2018 · 2018
Later among the works it cites.
Using perceptual and cognitive explanations for enhanced human-agent team performance, in: International Conference on Engineering Psychology and Cognitive Ergonomics, Springer. pp. 204–214
Neerincx, M.A., van der Waa, J., Kaptein, F., van Diggelen, J., 2018 · 2018
Later among the works it cites.
Hierarchical active inference: A theory of motivated control
Pezzulo, G., Rigoli, F., Friston, K.J., 2018 · 2018
Later among the works it cites.
Multi-goal reinforcement learning: Challenging robotics environments and request for research
Plappert, M., Andrychowicz, M., Ray, A., McGrew, B., Baker, B., Powell, G., Schneider, J., Tobin, J., Chociej, M., Welinder, P., Kumar, V., Zaremba, W., 2018 · 2018
Later among the works it cites.
Socially-aware reinforcement learning for personalized human-robot interaction, in: Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems, International Foundation for Autonomous Agents and Multiagent Systems. pp. 1775–1777
Ritschel, H., 2018 · 2018
Later among the works it cites.
Contrastive explanation for machine learning
Robeer, M.J., 2018 · 2018
Later among the works it cites.
Toward explainable multi-objective probabilistic planning, in: 2018 IEEE/ACM 4th International Workshop on Software Engineering for Smart Cyber-Physical Systems (SEsCPS), IEEE. pp. 19–25
Sukkerd, R., Simmons, R., Garlan, D., 2018 · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction (Second Edition)
Sutton, R.S., Barto, A.G., 2018 · 2018
Later among the works it cites.
Human-aligned artificial intelligence is a multiobjective problem
Vamplew, P., Dazeley, R., Foale, C., Firmin, S., Mummery, J., 2018 · 2018
Later among the works it cites.
Deep reinforcement learning and the deadly triad
Van Hasselt, H., Doron, Y., Strub, F., Hessel, M., Sonnerat, N., Modayil, J., 2018 · 2018
Later among the works it cites.
Programmatically interpretable reinforcement learning
Verma, A., Murali, V., Singh, R., Kohli, P., Chaudhuri, S., 2018 · 2018
Later among the works it cites.
Contrastive explanations for reinforcement learning in terms of expected consequences
van der Waa, J., van Diggelen, J., Bosch, K.v.d., Neerincx, M., 2018 · 2018
Later among the works it cites.
Equality of opportunity in classification: A causal approach
Zhang, J., Bareinboim, E., 2018 · 2018
Later among the works it cites.
Visual interpretability for deep learning: a survey
Zhang, Q.s., Zhu, S.C., 2018 · 2018
Later among the works it cites.
A survey on generative adversarial networks and their variants methods, in: Twelfth International Conference on Machine Vision (ICMV 2019), International Society for Optics and Photonics. p. 114333N
Aissa, F.B., Mejdoub, M., Zaied, M., 2020 · 2019
Later among the works it cites.
Mental models of mere mortals with explanations of reinforcement learning
Anderson, A.A., 2019 · 2019
Later among the works it cites.
Intelligible explanations in intelligent systems
Anjomshoae, S., Främling, K., 2019 · 2019
Later among the works it cites.
Explainable agents and robots: Results from a systematic literature review, in: Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems, International Foundation for Autonomous Agents and Multiagent Systems. pp. 1078–1088
Anjomshoae, S., Najjar, A., Calvaresi, D., Främling, K., 2019 · 2019
Later among the works it cites.
Dot-to-dot: Explainable hierarchical reinforcement learning for robotic manipulation
Beyret, B., Shafti, A., Faisal, A.A., 2019 · 2019
Later among the works it cites.
Planning and visualization for a smart meeting room assistant
Chakraborti, T., Fadnis, K.P., Talamadupula, K., Dholakia, M., Srivastava, B., Kephart, J.O., Bellamy, R.K., 2019 · 2019
Later among the works it cites.
Multi-agent deep reinforcement learning for large-scale traffic signal control
Chu, T., Wang, J., Codecà, L., Li, Z., 2019 · 2019
Later among the works it cites.
Memory-based explainable reinforcement learning, in: The 32nd Australasian Joint Conference on Artificial Intelligence (AusAI-19), pp. 66–77
Cruz, F., Dazeley, R., Vamplew, P., 2019 · 2019
Later among the works it cites.
On design and evaluation of human-centered explainable AI systems
Ehsan, U., 2019 · 2019
Later among the works it cites.
Automated rationale generation: a technique for explainable AI and its effects on human perceptions, in: Proceedings of the 24th International Conference on Intelligent User Interfaces, ACM. pp. 263–274
Ehsan, U., Tambwekar, P., Chan, L., Harrison, B., Riedl, M.O., 2019 · 2019
Later among the works it cites.
Tge-viz: Mixed initiative plan visualization
Gopalakrishnan, S., Kambhampati, S., 2019 · 2019
Later among the works it cites.
Explainable artificial intelligence (XAI)
Gunning, D., 2019 · 2019
Later among the works it cites.
Bonsai (AGi3)
Hammond, M., 2019 · 2019
Later among the works it cites.
Emotion regulation based on multi-objective weighted reinforcement learning for human-robot interaction, in: 2019 12th Asian Control Conference (ASCC), IEEE. pp. 1402–1406
Hao, M., Cao, W., Liu, Z., Wu, M., Yuan, Y., 2019 · 2019
Later among the works it cites.
Explainable AI planning (XAIP): Overview and the case of contrastive explanation, in: Reasoning Web. Explainable Artificial Intelligence. Springer, pp. 277–282
Hoffmann, J., Magazzeni, D., 2019 · 2019
Later among the works it cites.
A comprehensive survey of deep learning for image captioning
Hossain, M., Sohel, F., Shiratuddin, M.F., Laga, H., 2019 · 2019
Later among the works it cites.
Enabling robots to communicate their objectives
Huang, S.H., Held, D., Abbeel, P., Dragan, A.D., 2019 · 2019
Later among the works it cites.
Explainable reinforcement learning via reward decomposition, in: IJCAI/ECAI Workshop on Explainable Artificial Intelligence
Juozapaitis, Z., Koul, A., Fern, A., Erwig, M., Doshi-Velez, F., 2019 · 2019
Later among the works it cites.
Explaining sympathetic actions of rational agents, in: International Workshop on Explainable, Transparent Autonomous Agents and Multi-Agent Systems, Springer. pp. 59–76
Kampik, T., Nieves, J.C., Lindgren, H., 2019 · 2019
Later among the works it cites.
Reinforcement learning: By experimenting, computers are figuring out how to do things that no programmer could teach them
Knight, W., 2017 · 2019
Later among the works it cites.
Model-based contrastive explanations for explainable planning
Krarup, B., Cashmore, M., Magazzeni, D., Miller, T., 2019 · 2019
Later among the works it cites.
Explainable artificial intelligence applications in NLP, biomedical, and malware classification: A literature review, in: Intelligent Computing-Proceedings of the Computing Conference, Springer. pp. 1269–1292
Mathews, S.M., 2019 · 2019
Later among the works it cites.
How google’s AI viewed the move no human could understand
Metz, C., 2017a · 2019
Later among the works it cites.
In two moves, AlphaGo and Lee Sedol redefined the future
Metz, C., 2017b · 2019
Later among the works it cites.
Interpretable machine learning
Molnar, C., 2019 · 2019
Later among the works it cites.
Strategic tasks for explainable reinforcement learning, in: Proceedings of the AAAI Conference on Artificial Intelligence, pp. 10007–10008
Pocius, R., Neal, L., Fern, A., 2019 · 2019
Later among the works it cites.
Cxplain: Causal explanations for model interpretation under uncertainty, in: Advances in Neural Information Processing Systems, pp. 10220–10230
Schwab, P., Karlen, W., 2019 · 2019
Later among the works it cites.
Improving human-robot interaction through explainable reinforcement learning, in: 2019 14th ACM/IEEE International Conference on Human-Robot Interaction (HRI), IEEE. pp. 751–753
Tabrez, A., Hayes, B., 2019 · 2019
Later among the works it cites.
AGI Innovations (AGi3)
Voss, P., 2019 · 2019
Later among the works it cites.
Do you trust me?: Increasing user-trust by integrating virtual agents in explainable AI interaction design, in: Proceedings of the 19th ACM International Conference on Intelligent Virtual Agents, ACM. pp. 7–9
Weitz, K., Schiller, D., Schlagowski, R., Huber, T., André, E., 2019 · 2019
Later among the works it cites.
Scientific explanation
Woodward, J., 2017 · 2019
Later among the works it cites.
When agents talk back: Rebellious explanations
Wright, B., Roberts, M., Aha, D.W., Brumback, B., 2019 · 2019
Later among the works it cites.
An emotion-based approach to reinforcement learning reward design, in: 2019 IEEE 16th International Conference on Networking, Sensing and Control (ICNSC), IEEE. pp. 346–351
Yu, H., Yang, P., 2019 · 2019
Later among the works it cites.
Reinforcement learning interpretation methods: A survey
Alharin, A., Doan, T.N., Sartipi, M., 2020 · 2020
Later among the works it cites.
Moody learners - explaining competitive behaviour of reinforcement learning agents, in: Proceedings of the IEEE International Conference on Development and Learning (ICDL-EpiRob 2020)
Barros, P., Tanevska, A., Cruz, F., Sciutti, A., 2020 · 2020
Later among the works it cites.
Explanations for dynamic programming, in: International Symposium on Practical Aspects of Declarative Languages, Springer. pp. 179–195
Erwig, M., Kumar, P., Fern, A., 2020 · 2020
Later among the works it cites.
Marleme: A multi-agent reinforcement learning model extraction library, in: 2020 International Joint Conference on Neural Networks (IJCNN), IEEE. pp. 1–8
Kazhdan, D., Shams, Z., Liò, P., 2020 · 2020
Later among the works it cites.
Dictionary
Merriam-Webster, 2020 · 2020
Later among the works it cites.
Generative causal explanations of black-box classifiers
O’Shaughnessy, M., Canal, G., Connor, M., Rozell, C., Davenport, M., 2020 · 2020
Later among the works it cites.
Interestingness elements for explainable reinforcement learning: Understanding agents’ capabilities and limitations
Sequeira, P., Gervasio, M., 2020 · 2020
Later among the works it cites.
One explanation does not fit all
Sokol, K., Flach, P., 2020 · 2020
Later among the works it cites.
Potential-based multiobjective reinforcement learning approaches to low-impact agents for AI safety (submitted)
Vamplew, P., Foale, C., Dazeley, R., 2020 · 2020
Later among the works it cites.
Levels of explainable artificial intelligence for human-aligned conversational explanations
Dazeley, R., Vamplew, P., Foale, C., Young, C., Aryal, S., Cruz, F., 2021 · 2021
Closest in time.
A practical guide to multi-objective reinforcement learning and planning
Hayes, C.F., Rădulescu, R., Bargiacchi, E., Källström, J., Macfarlane, M., Reymond, M., Verstraeten, T., Zintgraf, L.M., Dazeley, R., Heintz, F., et al., 2021 · 2021
Closest in time.
Potential-based multiobjective reinforcement learning approaches to low-impact agents for ai safety
Vamplew, P., Foale, C., Dazeley, R., Bignold, A., 2021 · 2021
Closest in time.
Explainable embodied agents through social cues: A review
Wallkötter, S., Tulli, S., Castellano, G., Paiva, A., Chetouani, M., 2021 · 2021
Closest in time.
Explicable planning as minimizing distance from expected behavior, in: Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems, International Foundation for Autonomous Agents and Multiagent Systems. pp. 2075–2077
Kulkarni, A., Zha, Y., Chakraborti, T., Vadlamudi, S.G., Zhang, Y., Kambhampati, S., 2019 · 2077
Closest in time.
Toward robust policy summarization, in: Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems, International Foundation for Autonomous Agents and Multiagent Systems. pp. 2081–2083
Lage, I., Lifschitz, D., Doshi-Velez, F., Amir, O., 2019b · 2083
Closest in time.