Fetching the paper…
Reading the bibliography…
Despite the surprising power of many modern AI systems that often learn their own representations, there is significant discontent about their inscrutability and the attendant problems in their ability to interact with humans.
Garcez, A. d.; Gori, M.; Lamb, L. C.; Serafini, L.; Spranger, M.; and Tran, S. N. 2019 · 1905
Earlier work this paper cites.
Tractatus logico-philosophicus
Wittgenstein, L. 1922 · 1922
Earlier work this paper cites.
Programs with Common Sense, proceedings of the Teddington Conference on the Mechanization of Thought Processes, 75-91
McCarthy, J. 1959 · 1959
Earlier work this paper cites.
Intelligent tutoring systems
Anderson, J. R.; Boyle, C. F.; and Reiser, B. J. 1985 · 1985
Earlier work this paper cites.
PDDL-the planning domain definition language
McDermott, D.; Ghallab, M.; Howe, A.; Knoblock, C.; Ram, A.; Veloso, M.; Weld, D.; and Wilkins, D. 1998 · 1998
Earlier work this paper cites.
Reinforcement learning with human teachers: Evidence of feedback and guidance with implications for learning performance
Thomaz, A. L.; Breazeal, C.; et al. 2006 · 2006
Earlier work this paper cites.
Causality
Pearl, J. 2009 · 2009
Earlier work this paper cites.
Thinking, fast and slow
Kahneman, D. 2011 · 2011
Earlier work this paper cites.
Learning from explanations using sentiment and advice in RL
Krening, S.; Harrison, B.; Feigh, K. M.; Isbell, C. L.; Riedl, M.; and Thomaz, A. 2016 · 2016
Earlier work this paper cites.
Plan Explanations as Model Reconciliation: Moving Beyond Explanation as Soliloquy
Chakraborti, T.; Sreedharan, S.; Zhang, Y.; and Kambhampati, S. 2017 · 2017
Earlier work this paper cites.
Deep Reinforcement Learning from Human Preferences
Christiano, P. F.; Leike, J.; Brown, T. B.; Martic, M.; Legg, S.; and Amodei, D. 2017 · 2017
Cited alongside, same era.
The off-switch game
Hadfield-Menell, D.; Dragan, A.; Abbeel, P.; and Russell, S. 2017 · 2017
Cited alongside, same era.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
Krishna, R.; Zhu, Y.; Groth, O.; Johnson, J.; Hata, K.; Kravitz, J.; Chen, S.; Kalantidis, Y.; Li, L.-J.; Shamma, D. A.; et al. 2017 · 2017
Cited alongside, same era.
Visualizing and Understanding Atari Agents
Greydanus, S.; Koul, A.; Dodge, J.; and Fern, A. 2018 · 2018
Cited alongside, same era.
Using reward machines for high-level task specification and decomposition in reinforcement learning
Icarte, R. T.; Klassen, T.; Valenzano, R.; and McIlraith, S. 2018 · 2018
Cited alongside, same era.
From skills to symbols: Learning symbolic representations for abstract high-level planning
Challenges of Human-Aware AI Systems AAAI Presidential Address
Kambhampati, S. 2020 · 2020
Later among the works it cites.
Bridging the Gap: Providing Post-Hoc Symbolic Explanations for Sequential Decision-Making Problems with Inscrutable Representations
Sreedharan, S.; Soni, U.; Verma, M.; Srivastava, S.; and Kambhampati, S. 2020 · 2020
Later among the works it cites.
Atari-head: Atari human eye-tracking and demonstration dataset
Zhang, R.; Walshe, C.; Liu, Z.; Guan, L.; Muller, K.; Whritner, J.; Zhang, L.; Hayhoe, M.; and Ballard, D. 2020 · 2020
Later among the works it cites.
Widening the Pipeline in Human-Guided Reinforcement Learning with Explanation and Context-Aware Data Augmentation
Guan, L.; Verma, M.; Guo, S.; Zhang, R.; and Kambhampati, S. 2021 · 2021
Closest in time.
Polanyi’s revenge and AI’s new romance with tacit knowledge
Kambhampati, S. 2021 · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Konidaris, G.; Kaelbling, L. P.; and Lozano-Perez, T. 2018 · 2018
Cited alongside, same era.
Learning first-order symbolic representations for planning from the structure of the state space
Bonet, B.; and Geffner, H. 2019 · 2019
Cited alongside, same era.
Foundations for restraining bolts: Reinforcement learning with LTLf/LDLf restraining specifications
De Giacomo, G.; Iocchi, L.; Favorito, M.; and Patrizi, F. 2019 · 2019
Cited alongside, same era.
Neuro-symbolic= neural+ logical+ probabilistic
De Raedt, L.; Manhaeve, R.; Dumancic, S.; Demeester, T.; and Kimmig, A. 2019 · 2019
Cited alongside, same era.
Towards automatic concept-based explanations
Ghorbani, A.; Wexler, J.; Zou, J. Y.; and Kim, B. 2019 · 2019
Cited alongside, same era.
McGrath, T.; Kapishnikov, A.; Tomašev, N.; Pearce, A.; Hassabis, D.; Kim, B.; Paquet, U.; and Kramnik, V. 2021 · 2021
Closest in time.
Foundations of explanations as model reconciliation
Sreedharan, S.; Chakraborti, T.; and Kambhampati, S. 2021 · 2021
Closest in time.
Explainable Human-AI Interaction: A Planning Perspective
Sreedharan, S.; Kulkarni, A.; and Kambhampati, S. 2021 · 2021
Closest in time.
Using state abstractions to compute personalized contrastive explanations for AI agent behavior
Sreedharan, S.; Srivastava, S.; and Kambhampati, S. 2021 · 2021
Closest in time.
Learning from Ambiguous Demonstrations with Self-Explanation Guided Reinforcement Learning
Zha, Y.; Guan, L.; and Kambhampati, S. 2021 · 2021
Closest in time.