Fetching the paper…
Reading the bibliography…
Deep reinforcement learning agents are prone to goal misalignments.
Classification And Regression Trees
Leo Breiman, Jerome Friedman, R.A. Olshen, and Charles J. Stone · 1984
Earlier work this paper cites.
A system for induction of oblique decision trees
Sreerama K Murthy, Simon Kasif, and Steven Salzberg · 1994
Earlier work this paper cites.
Lookahead and pathology in decision tree induction
Sreerama Murthy and Steven Salzberg · 1995
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey J. Gordon, and J. Andrew Bagnell · 2010
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents (extended abstract)
Marc G. Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller · 2013
Earlier work this paper cites.
Markov decision processes: discrete stochastic dynamic programming
Martin L Puterman · 2014
Earlier work this paper cites.
Introduction to evolutionary computing
Agoston E Eiben and James E Smith · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin A. Riedmiller, Andreas Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Earlier work this paper cites.
Trust region policy optimization
John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz · 2015
Earlier work this paper cites.
The mythos of model interpretability
Zachary Chase Lipton · 2016
Earlier work this paper cites.
Deep reinforcement learning with double q-learning
Hado van Hasselt, Arthur Guez, and David Silver · 2016
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Verifiable reinforcement learning via policy extraction
Osbert Bastani, Yewen Pu, and Armando Solar-Lezama · 2018
Earlier work this paper cites.
A survey of methods for explaining black box models
Riccardo Guidotti, Anna Monreale, Salvatore Ruggieri, Franco Turini, Fosca Giannotti, and Dino Pedreschi · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Revisiting the arcade learning environment: Evaluation protocols and open problems for general agents (extended abstract)
Marlos C. Machado, Marc G. Bellemare, Erik Talvitie, Joel Veness, Matthew J. Hausknecht, and Michael Bowling · 2018
Cited alongside, same era.
Computational Complexity Analysis of Decision Tree Algorithms: 38th SGAI International Conference on Artificial Intelligence, AI 2018, Cambridge, UK, December 11–13, 2018, Proceedings , pages 191–197
Habiba Sani, Ci Lei, and Daniel Neagu · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Cited alongside, same era.
Programmatically interpretable reinforcement learning
Abhinav Verma, Vijayaraghavan Murali, Rishabh Singh, Pushmeet Kohli, and Swarat Chaudhuri · 2018
Cited alongside, same era.
Efficient decompositional rule extraction for deep neural networks
Mateo Espinosa Zarlenga, Zohreh Shams, and Mateja Jamnik · 2021
Later among the works it cites.
Murtree: Optimal decision trees via dynamic programming and search
Emir Demirovic, Anna Lukina, Emmanuel Hebrard, Jeffrey Chan, James Bailey, Christopher Leckie, Kotagiri Ramamohanarao, and Peter J. Stuckey · 2022
Later among the works it cites.
Goal misgeneralization in deep reinforcement learning
Lauro Langosco di Langosco, Jack Koch, Lee D. Sharkey, Jacob Pfau, and David Krueger · 2022
Later among the works it cites.
Quant-BnB: A scalable branch-and-bound method for optimal decision trees with continuous features
Rahul Mazumder, Xiang Meng, and Haoyue Wang · 2022
Later among the works it cites.
A survey of explainable reinforcement learning
Stephanie Milani, Nicholay Topin, Manuela M. Veloso, and Fei Fang · 2022
Later among the works it cites.
Programmatic reinforcement learning without oracles
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Definitions, methods, and applications in interpretable machine learning
W. James Murdoch, Chandan Singh, Karl Kumbier, Reza Abbasi-Asl, and Bin Yu · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala · 2019
Cited alongside, same era.
Information Fusion , 2020
Explainable artificial intelligence (xai): Concepts, taxonomies, opportunities and challenges toward responsible ai · 2020
Cited alongside, same era.
Model interpretability through the lens of computational complexity
Pablo Barceló, Mikaël Monet, Jorge Pérez, and Bernardo Subercaseaux · 2020
Cited alongside, same era.
Model interpretability through the lens of computational complexity
Pablo Barceló, Mikaël Monet, Jorge Pérez, and Bernardo Subercaseaux · 2020
Cited alongside, same era.
SPACE: unsupervised object-oriented scene representation via spatial attention and decomposition
Zhixuan Lin, Yi-Fu Wu, Skand Vishwanath Peri, Weihao Sun, Gautam Singh, Fei Deng, Jindong Jiang, and Sungjin Ahn · 2020
Cited alongside, same era.
Rl baselines3 zoo, 2020
Antonin Raffin · 2020
Cited alongside, same era.
Wenjie Qiu and He Zhu · 2022
Later among the works it cites.
Explainable deep learning: A field guide for the uninitiated
Gabrielle Ras, Ning Xie, Marcel van Gerven, and Derek Doran · 2022
Later among the works it cites.
Explainability via causal self-talk
Nicholas A. Roy, Junkyung Kim, and Neil C. Rabinowitz · 2022
Later among the works it cites.
Learning crop management by reinforcement: gym-dssat
Romain Gautron, Emilio J Padrón, Philippe Preux, Julien Bigot, Odalric-Ambrym Maillard, Gerrit Hoogenboom, and Julien Teigny · 2023
Later among the works it cites.
Explainable AI (XAI): A systematic meta-survey of current challenges and future opportunities
Waddah Saeed and Christian W. Omlin · 2023
Later among the works it cites.
Gymnasium, 2023
Mark Towers, Jordan K. Terry, Ariel Kwiatkowski, John U. Balis, Gianluca de Cola, Tristan Deleu, Manuel Goulão, Andreas Kallinteris, Arjun KG, Markus Krimmel, Rodrigo Perez-Vicente, Andrea Pierré, Sander Schulhoff, Jun Jet Tai, Andrew Tan Jin Shen, and Omar G. Younis · 2023
Later among the works it cites.
Fast segment anything
Xu Zhao, Wenchao Ding, Yongqi An, Yinglong Du, Tao Yu, Min Li, Ming Tang, and Jinqiao Wang · 2023
Later among the works it cites.
Assessing the interpretability of programmatic policies with large language models, 2024
Zahra Bashir, Michael Bowling, and Levi H. S. Lelis · 2024
Closest in time.
Crossq: Batch normalization in deep reinforcement learning for greater sample efficiency and simplicity
Aditya Bhatt, Daniel Palenicek, Boris Belousov, Max Argus, Artemij Amiranashvili, Thomas Brox, and Jan Peters · 2024
Closest in time.
Interpretable concept bottlenecks to align reinforcement learning agents
Quentin Delfosse, Sebastian Sztwiertnia, Wolfgang Stammer, Mark Rothermel, and Kristian Kersting · 2024
Closest in time.
Interpretable decision tree search as a markov decision process, 2024
Hector Kohler, Riad Akrour, and Philippe Preux · 2024
Closest in time.
Insight: End-to-end neuro-symbolic visual reinforcement learning with language explanations
Lirui Luo, Guoxi Zhang, Hongming Xu, Yaodong Yang, Cong Fang, and Qing Li · 2024
Closest in time.
Pix2code: Learning to compose neural visual concepts as programs
Antonia Wüst, Wolfgang Stammer, Quentin Delfosse, Devendra Singh Dhami, and Kristian Kersting · 2024
Closest in time.