Fetching the paper…
Reading the bibliography…
Deep Reinforcement Learning (DRL) has achieved remarkable success in sequential decision-making problems.
S. Collins, A. Ruina, R. Tedrake, and M. Wisse, “Efficient bipedal robots based on passive-dynamic walkers,” Science , vol. 307, no. 5712, pp. 1082–1085, 2005
2005
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2011, pp. 627–635
2011
Earlier work this paper cites.
O. Bastani, Y. Ioannou, L. Lampropoulos, D. Vytiniotis, A. Nori, and A. Criminisi, “Measuring neural net robustness with constraints,” Advances in neural information processing systems , vol. 29, 2016
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of go with deep neural networks and tree search,” nature , vol. 529, no. 7587, pp. 484–489, 2016
2016
Earlier work this paper cites.
M. Lanctot, V. Zambaldi, A. Gruslys, A. Lazaridou, K. Tuyls, J. Pérolat, D. Silver, and T. Graepel, “A unified game-theoretic approach to multiagent reinforcement learning,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 30, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne, “Imitation learning: A survey of learning methods,” ACM Computing Surveys (CSUR) , vol. 50, no. 2, pp. 1–35, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
O. Bastani, Y. Pu, and A. Solar-Lezama, “Verifiable reinforcement learning via policy extraction,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 31, 2018
2018
Earlier work this paper cites.
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev et al. , “Grandmaster level in starcraft ii using multi-agent reinforcement learning,” Nature , vol. 575, no. 7782, pp. 350–354, 2019
2019
Earlier work this paper cites.
T. Johannink, S. Bahl, A. Nair, J. Luo, A. Kumar, M. Loskyll, J. A. Ojea, E. Solowjow, and S. Levine, “Residual reinforcement learning for robot control,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 6023–6029
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Z. Liu, M. Sun, T. Zhou, G. Huang, and T. Darrell, “Rethinking the value of network pruning,” International Conference on Learning Representations (ICLR) , 2019
2019
Cited alongside, same era.
G. Liu, O. Schulte, W. Zhu, and Q. Li, “Toward interpretable deep reinforcement learning with linear model u-trees,” in Machine Learning and Knowledge Discovery in Databases: European Conference, ECML PKDD 2018, Dublin, Ireland, September 10–14, 2018, Proceedings, Part II 18 . Springer, 2019, pp. 414–429
2019
Cited alongside, same era.
Y. Coppens, K. Efthymiadis, T. Lenaerts, A. Nowé, T. Miller, R. Weber, and D. Magazzeni, “Distilling deep reinforcement learning policies in soft decision trees,” in Proceedings of the IJCAI 2019 workshop on explainable artificial intelligence , 2019, pp. 1–6
2019
Cited alongside, same era.
A. Rosenfeld and A. Richardson, “Explainability in human–agent systems,” Autonomous Agents and Multi-Agent Systems , vol. 33, pp. 673–705, 2019
2019
Cited alongside, same era.
W. Guo, X. Wu, U. Khan, and X. Xing, “Edge: Explaining deep reinforcement learning policies,” Advances in Neural Information Processing Systems , vol. 34, pp. 12 222–12 236, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
A. Romdhana, A. Merlo, M. Ceccato, and P. Tonella, “Deep reinforcement learning for black-box testing of android apps,” ACM Transactions on Software Engineering and Methodology (TOSEM) , vol. 31, no. 4, pp. 1–29, 2022
2022
Later among the works it cites.
D. Minh, H. X. Wang, Y. F. Li, and T. N. Nguyen, “Explainable artificial intelligence: a comprehensive review,” Artificial Intelligence Review , pp. 1–66, 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Molnar, Interpretable machine learning . Lulu. com, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
S. K. S. Ghasemipour, R. Zemel, and S. Gu, “A divergence minimization perspective on imitation learning methods,” in Conference on Robot Learning . PMLR, 2020, pp. 1259–1277
2020
Cited alongside, same era.
H. Cha, J. Park, H. Kim, M. Bennis, and S.-L. Kim, “Proxy experience replay: Federated distillation for distributed reinforcement learning,” IEEE Intelligent Systems , vol. 35, no. 4, pp. 94–101, 2020
2020
Cited alongside, same era.
E. Tjoa and C. Guan, “A survey on explainable artificial intelligence (xai): Toward medical xai,” IEEE transactions on neural networks and learning systems , vol. 32, no. 11, pp. 4793–4813, 2020
2020
Cited alongside, same era.
E. Puiutta and E. M. Veith, “Explainable reinforcement learning: A survey,” in International cross-domain conference for machine learning and knowledge extraction . Springer, 2020, pp. 77–95
2020
Cited alongside, same era.
A. Silva, M. Gombolay, T. Killian, I. Jimenez, and S.-H. Son, “Optimization methods for interpretable differentiable decision trees applied to reinforcement learning,” in International conference on artificial intelligence and statistics . PMLR, 2020, pp. 1855–1865
2020
Cited alongside, same era.
K. Kurach, A. Raichuk, P. Stańczyk, M. Zając, O. Bachem, L. Espeholt, C. Riquelme, D. Vincent, M. Michalski, O. Bousquet et al. , “Google research football: A novel reinforcement learning environment,” in Proceedings of the AAAI conference on artificial intelligence , vol. 34, no. 04, 2020, pp. 4501–4510
2020
Cited alongside, same era.
H. W. Loh, C. P. Ooi, S. Seoni, P. D. Barua, F. Molinari, and U. R. Acharya, “Application of explainable artificial intelligence for healthcare: A systematic review of the last decade (2011–2022),” Computer Methods and Programs in Biomedicine , p. 107161, 2022
2022
Later among the works it cites.
M. Vasić, A. Petrović, K. Wang, M. Nikolić, R. Singh, and S. Khurshid, “Moet: Mixture of expert trees and its application to verifiable reinforcement learning,” Neural Networks , vol. 151, pp. 34–47, 2022
2022
Later among the works it cites.
Y. You, J. Sun, Y. Guo, Y. Tan, and J. Jiang, “Interpretability and accuracy trade-off in the modeling of belief rule-based systems,” Knowledge-Based Systems , vol. 236, p. 107491, 2022
2022
Later among the works it cites.
S. Milani, Z. Zhang, N. Topin, Z. R. Shi, C. Kamhoua, E. E. Papalexakis, and F. Fang, “Maviper: Learning decision tree policies for interpretable multi-agent reinforcement learning,” in Joint European Conference on Machine Learning and Knowledge Discovery in Databases . Springer, 2022, pp. 251–266
2022
Later among the works it cites.
C. Li, P. Zheng, Y. Yin, B. Wang, and L. Wang, “Deep reinforcement learning in smart manufacturing: A review and prospects,” CIRP Journal of Manufacturing Science and Technology , vol. 40, pp. 75–101, 2023
2023
Closest in time.
Y. Guo, J. Campbell, S. Stepputtis, R. Li, D. Hughes, F. Fang, and K. Sycara, “Explainable action advising for multi-agent reinforcement learning,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 5515–5521
2023
Closest in time.
Y.-C. Wang, T. Chen, and M.-C. Chiu, “An explainable deep-learning approach for job cycle time prediction,” Decision Analytics Journal , vol. 6, p. 100153, 2023
2023
Closest in time.
X. Liu, S. Liu, B. An, Y. Gao, S. Yang, and W. Li, “Effective interpretable policy distillation via critical experience point identification,” IEEE Intelligent Systems , 2023
2023
Closest in time.