Fetching the paper…
Reading the bibliography…
Intelligent agents such as robots are increasingly deployed in real-world, safety-critical settings.
“Language models are few-shot learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“The need for user models in generating expert system explanations”
Robert Kass and Tim Finin · 1988
Earlier work this paper cites.
“Agents that Learn to Explain Themselves.”
W Johnson · 1994
Earlier work this paper cites.
“Do deep nets really need to be deep?”
Jimmy Ba and Rich Caruana · 2014
Earlier work this paper cites.
“Distilling the knowledge in a neural network”
Geoffrey Hinton, Oriol Vinyals and Jeff Dean · 2015
Earlier work this paper cites.
“Natural language generation enhances human decision-making with uncertain information”
Dimitra Gkatzia, Oliver Lemon and Verena Rieser · 2016
Earlier work this paper cites.
“” Why should i trust you?” Explaining the predictions of any classifier”
Marco Ribeiro, Sameer Singh and Carlos Guestrin · 2016
Earlier work this paper cites.
“Survey of machine learning algorithms for disease diagnostic”
Meherwar Fatima and Maruf Pasha · 2017
Earlier work this paper cites.
“An exploratory study on the benefits of using natural language for explaining fuzzy rule-based systems”
Jose Alonso, Alejandro Ramos-Soto, Ehud Reiter and Kees van Deemter · 2017
Earlier work this paper cites.
“Improving robot controller transparency through autonomous policy explanation”
Bradley Hayes and Julie Shah · 2017
Earlier work this paper cites.
“Hierarchical and interpretable skill acquisition in multi-task reinforcement learning”
Tianmin Shu, Caiming Xiong and Richard Socher · 2017
Earlier work this paper cites.
“Distilling a neural network into a soft decision tree”
Nicholas Frosst and Geoffrey Hinton · 2017
Earlier work this paper cites.
“Interpretability via model extraction”
Osbert Bastani, Carolyn Kim and Hamsa Bastani · 2017
Earlier work this paper cites.
“Programmatically interpretable reinforcement learning”
Abhinav Verma et al · 2018
Earlier work this paper cites.
“e-snli: Natural language inference with natural language explanations”
Oana-Maria Camburu, Tim Rocktäschel, Thomas Lukasiewicz and Phil Blunsom · 2018
Earlier work this paper cites.
“Verifiable reinforcement learning via policy extraction”
Osbert Bastani, Yewen Pu and Armando Solar-Lezama · 2018
Earlier work this paper cites.
“Summarizing agent strategies”
Ofra Amir, Finale Doshi-Velez and David Sarne · 2019
Earlier work this paper cites.
“XAI—Explainable artificial intelligence”
David Gunning et al · 2019
Earlier work this paper cites.
“Verbal explanations for deep reinforcement learning neural networks with attention on extracted features”
Xinzhi Wang et al · 2019
Cited alongside, same era.
“Generating justifications for norm-related agent decisions”
Daniel Kasenberg et al · 2019
Cited alongside, same era.
“Automated rationale generation: a technique for explainable AI and its effects on human perceptions”
Upol Ehsan et al · 2019
Cited alongside, same era.
“Explanation in artificial intelligence: Insights from the social sciences”
Tim Miller · 2019
Cited alongside, same era.
“Toward interpretable deep reinforcement learning with linear model u-trees”
Guiliang Liu, Oliver Schulte, Wang Zhu and Qingcan Li · 2019
Cited alongside, same era.
“Few-shot self-rationalization with natural language prompts”
Ana Marasović, Iz Beltagy, Doug Downey and Matthew Peters · 2021
Later among the works it cites.
“To what extent do human explanations of model behavior align with actual model behavior?”
Grusha Prasad et al · 2021
Later among the works it cites.
“Why? why not? when? visual explanations of agent behaviour in reinforcement learning”
Aditi Mishra, Utkarsh Soni, Jinbin Huang and Chris Bryan · 2022
Later among the works it cites.
“Explanations from large language models make small reasoners better”
Shiyang Li et al · 2022
Later among the works it cites.
“A System for Imitation Learning of Contact-Rich Bimanual Manipulation Policies”
Simon Stepputtis, Maryam Bandari, Stefan Schaal and Heni Amor · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mengnan Du, Ninghao Liu and Xia Hu · 2019
Cited alongside, same era.
“Explain Yourself! Leveraging Language Models for Commonsense Reasoning”
Nazneen Rajani, Bryan McCann, Caiming Xiong and Richard Socher · 2019
Cited alongside, same era.
“Scalability in perception for autonomous driving: Waymo open dataset”
Pei Sun et al · 2020
Cited alongside, same era.
“Machine learning-based traffic prediction models for intelligent transportation systems”
Azzedine Boukerche and Jiahao Wang · 2020
Cited alongside, same era.
“Towards harnessing natural language generation to explain black-box models”
Ettore Mariotti, Jose Alonso and Albert Gatt · 2020
Cited alongside, same era.
“Explainable reinforcement learning: A survey”
Erika Puiutta and Eric Veith · 2020
Cited alongside, same era.
“Learning whole-body human-robot haptic interaction in social contexts”
Joseph Campbell and Katsu Yamane · 2020
Cited alongside, same era.
“Knowledge-Grounded Self-Rationalization via Extractive and Natural Language Explanations”, 2022
2022
Later among the works it cites.
“Learning modular language-conditioned robot policies through attention”
Yifan Zhou et al · 2023
Closest in time.
“Large-Scale Package Manipulation via Learned Metrics of Pick Success”
Shuai Li et al · 2023
Closest in time.
“Concept learning for interpretable multi-agent reinforcement learning”
Renos Zabounidis et al · 2023
Closest in time.
“Introspective Action Advising for Interpretable Transfer Learning”
Joseph Campbell et al · 2023
Closest in time.
“A novel policy-graph approach with natural language and counterfactual abstractions for explaining reinforcement learning agents”
Tongtong Liu et al · 2023
Closest in time.
“Sources of Hallucination by Large Language Models on Inference Tasks”
Nick McKenna et al · 2023
Closest in time.
“Sample-Efficient Learning of Novel Visual Concepts”
Sarthak Bhagat, Simon Stepputtis, Joseph Campbell and Katia Sycara · 2023
Closest in time.
“Explainable action advising for multi-agent reinforcement learning”
Yue Guo et al · 2023
Closest in time.
“Thought Cloning: Learning to Think while Acting by Imitating Human Thinking”
Shengran Hu and Jeff Clune · 2023
Closest in time.
“Language models can explain neurons in language models”
Steven Bills et al · 2023
Closest in time.
OpenAI, 2023
“GPT-4” · 2023
Closest in time.