Fetching the paper…
Reading the bibliography…
Human environments are often regulated by explicit and complex rulesets.
The nature of explanation
Peter Achinstein · 1983
Earlier work this paper cites.
Occam’s razor
Anselm Blumer, Andrzej Ehrenfeucht, David Haussler, and Manfred K Warmuth · 1987
Earlier work this paper cites.
Induction: Processes of inference, learning, and discovery
John H Holland, Keith J Holyoak, Richard E Nisbett, and Paul R Thagard · 1989
Earlier work this paper cites.
Building explanations from rules and structured cases
L.Karl Branting · 1991
Earlier work this paper cites.
Case-Based Reasoning: Foundational Issues, Methodological Variations, and System Approaches
Agnar Aamodt and Enric Plaza · 1994
Earlier work this paper cites.
Enhancing the explanatory power of usability heuristics
Jakob Nielsen · 1994
Earlier work this paper cites.
Explanation-based learning and reinforcement learning: A unified view
Thomas G Dietterich and Nicholas S Flann · 1997
Earlier work this paper cites.
Driver Braking Performance in Stopping Sight Distance Situations
Daniel B Fambro, Rodger J Koppa, Dale L Picha, and Kay Fitzpatrick · 2000
Earlier work this paper cites.
Theories of explanation. the internet encyclopedia of philosophy, 2005
GR Mayes · 2005
Earlier work this paper cites.
A comprehensive survey on safe reinforcement learning
Javier Garcıa and Fernando Fernández · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Earlier work this paper cites.
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2015
Earlier work this paper cites.
Safe Model-based Reinforcement Learning with Stability Guarantees
Felix Berkenkamp, Matteo Turchetta, Angela Schoellig, and Andreas Krause · 2017
Cited alongside, same era.
Acceleration-Deceleration Behaviour of Various Vehicle Types
P S Bokare and A K Maurya · 2017
Cited alongside, same era.
Imitation learning: A survey of learning methods
Ahmed Hussein, Mohamed Medhat Gaber, Eyad Elyan, and Chrisina Jayne · 2017
Cited alongside, same era.
Ray rllib: A composable and scalable reinforcement learning library
Eric Liang, Richard Liaw, Robert Nishihara, Philipp Moritz, Roy Fox, Joseph Gonzalez, Ken Goldberg, and Ion Stoica · 2017
Cited alongside, same era.
Knowledge transfer for deep reinforcement learning with hierarchical experience replay
Haiyan Yin and Sinno Pan · 2017
Cited alongside, same era.
Urban Driving with Multi-Objective Deep Reinforcement Learning
Changjian Li and Krzysztof Czarnecki · 2019
Later among the works it cites.
Combining experience replay with exploration by random network distillation
Francesco Sovrano · 2019
Later among the works it cites.
Experience replay optimization
Daochen Zha, Kwei-Herng Lai, Kaixiong Zhou, and Xia Hu · 2019
Later among the works it cites.
Explainable Artificial Intelligence (XAI): Concepts, taxonomies, opportunities and challenges toward responsible AI
Alejandro Barredo Arrieta, Natalia Díaz-Rodríguez, Javier Del Ser, Adrien Bennetot, Siham Tabik, Alberto Barbado, Salvador Garcia, Sergio Gil-Lopez, Daniel Molina, Richard Benjamins, Raja Chatila, and Francisco Herrera · 2020
Later among the works it cites.
Culture-Based Explainable Human-Agent Deconfliction
Alex Raymond, Hatice Gunes, and Amanda Prorok · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yinlam Chow, Ofir Nachum, Edgar Duenez-Guzman, and Mohammad Ghavamzadeh · 2018
Cited alongside, same era.
Addressing Function Approximation Error in Actor-Critic Methods
Scott Fujimoto, Herke van Hoof, and David Meger · 2018
Cited alongside, same era.
Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Self-paced prioritized curriculum learning with coverage penalty in deep reinforcement learning
Zhipeng Ren, Daoyi Dong, Huaxiong Li, and Chunlin Chen · 2018
Cited alongside, same era.
Simplicity and complexity preferences in causal explanation: An opponent heuristic account
Samuel GB Johnson, JJ Valenti, and Frank C Keil · 2019
Cited alongside, same era.
Model-based reinforcement learning for atari
Lukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski, Roy H Campbell, Konrad Czechowski, Dumitru Erhan, Chelsea Finn, Piotr Kozakowski, Sergey Levine, and others · 2019
Cited alongside, same era.
Safe Reinforcement Learning with Policy-Guided Planning for Autonomous Driving
Jikun Rong and Nan Luan · 2020
Later among the works it cites.
Attentive experience replay
Peiquan Sun, Wengang Zhou, and Houqiang Li · 2020
Later among the works it cites.
Deep Reinforcement Learning for Autonomous Driving: A Survey
B Ravi Kiran, Ibrahim Sobh, Victor Talpaert, Patrick Mannion, Ahmad A Al Sallab, Senthil Yogamani, and Patrick Pérez · 2021
Closest in time.
Revisiting Prioritized Experience Replay: A Value Perspective
Ang A Li, Zongqing Lu, and Chenglin Miao · 2021
Closest in time.
Agree to Disagree: Subjective Fairness in Privacy-Restricted Decentralised Conflict Resolution
Alex Raymond, Matthew Malencia, Guilherme Paulino-Passos, and Amanda Prorok · 2021
Closest in time.
From philosophy to interfaces: an explanatory method and a tool based on achinstein’s theory of explanation
Francesco Sovrano and Fabio Vitali · 2021
Closest in time.