Fetching the paper…
Reading the bibliography…
Despite numerous successes in Deep Reinforcement Learning (DRL), the learned policies are not interpretable.
Fikes, R.E., Nilsson, N.J.: Strips: A new approach to the application of theorem proving to problem solving. Artificial Intelligence 2
1971
Earlier work this paper cites.
Summers, P.D.: A methodology for lisp program construction from examples. J. ACM 24
1977
Earlier work this paper cites.
Biermann, A.W.: The inference of regular lisp programs from examples. IEEE Transactions on Systems, Man, and Cybernetics 8
1978
Earlier work this paper cites.
Lloyd, J.W.: Foundations of Logic Programming. Springer-Verlag, Berlin, Heidelberg (1984)
1984
Earlier work this paper cites.
Khoshafian, S.N., Copeland, G.P.: Object identity. ACM SIGPLAN Notices 21
1986
Earlier work this paper cites.
Quinlan, J.R.: Learning logical definitions from relations. Mach. Learn. 5
1990
Earlier work this paper cites.
Williams, R.J.: Simple statistical gradient following algorithms for connectionist reinforcement learning. Machine Learning 8
1992
Earlier work this paper cites.
Muggleton, S., De Raedt, L.: Inductive logic programming: Theory and methods. Journal Of Logic Programming 19
1994
Earlier work this paper cites.
Dzeroski, S., Raedt, L.D., Blockeel, H.: Relational reinforcement learning. In: Proceedings of the Fifteenth International Conference on Machine Learning. p. 136–143. ICML ’98, Morgan Kaufmann Publishers Inc., San Francisco, CA, USA (1998)
1998
Earlier work this paper cites.
Gelfond, M., Lifschitz, V.: Action languages. Electronic Transactions on Artificial Intelligence 3
1998
Earlier work this paper cites.
Ghallab, M., Howe, A., Knoblock, C., Mcdermott, D., Ram, A., Veloso, M., Weld, D., Wilkins, D.: PDDL—The Planning Domain Definition Language (1998), http://citeseerx.ist.psu.edu/viewdoc/summary?doi=10.1.1.37.212
1998
Earlier work this paper cites.
Esteva, F., Godo, L.: Monoidal t-norm based logic:towards a logic for left-continuous t-norms. Fuzzy Sets and Systems 124
2001
Earlier work this paper cites.
Driessens, K., Džeroski, S.: Integrating guidance into relational reinforcement learning. Machine Learning 57
2004
Earlier work this paper cites.
Kersting, K., Otterlo, M.V., De Raedt, L.: Bellman goes relational. In: Proceedings of the twenty-first international conference on Machine learning. p. 59 (2004)
2004
Earlier work this paper cites.
Lee, J.D., See, K.A.: Trust in automation: Designing for appropriate reliance. Human factors 46 1
2004
Earlier work this paper cites.
Nau, D., Ghallab, M., Traverso, P.: Automated Planning: Theory & Practice. Morgan Kaufmann Publishers Inc., San Francisco, CA, USA (2004)
2004
Earlier work this paper cites.
Tadepalli, P., Givan, R., Driessens, K.: Relational reinforcement learning: An overview. In: In: Proceedings of the ICML’04 Workshop on Relational Reinforcement Learning (2004)
2004
Earlier work this paper cites.
Ernst, D., Geurts, P., Wehenkel, L.: Tree-based batch mode reinforcement learning. Journal of Machine Learning Research 6
2005
Earlier work this paper cites.
Xu, R., Wunsch, D.: Survey of clustering algorithms. IEEE Transactions on Neural Networks 16
2005
Earlier work this paper cites.
De Raedt, L.: Logical and relational learning. In: Zaverucha, G., da Costa, A.L. (eds.) Advances in Artificial Intelligence - SBIA 2008. pp. 1–1. Springer Berlin Heidelberg, Berlin, Heidelberg (2008)
2008
Cited alongside, same era.
Kersting, K., Driessens, K.: Non-parametric policy gradients: A unified treatment of propositional and relational domains. In: Proceedings of the 25th International Conference on Machine learning. pp. 456–463 (2008)
2008
Cited alongside, same era.
Solar-Lezama, A.: Program Synthesis by Sketching. Ph.D. thesis, University of California at Berkeley, USA (2008), https://dl.acm.org/doi/10.5555/1714168 , aAI3353225
2008
Cited alongside, same era.
Gulwani, S., Harris, W.R., Singh, R.: Spreadsheet data manipulation using examples. Commun. ACM 55
2012
Cited alongside, same era.
Evans, R., Grefenstette, E.: Learning explanatory rules from noisy data. J. Artif. Int. Res. 61
2018
Later among the works it cites.
2018
Later among the works it cites.
Iyer, R., Li, Y., Li, H., Lewis, M., Sundar, R., Sycara, K.: Transparency and explanation in deep reinforcement learning neural networks. In: Proceedings of the 2018 AAAI/ACM Conference on AI, Ethics, and Society. p. 144–150. AIES ’18, Association for Computing Machinery, New York, NY, USA (2018), https://doi.org/10.1145/3278721.3278776
2018
Later among the works it cites.
Liu, G., Schulte, O., Zhu, W., Li, Q.: Toward interpretable deep reinforcement learning with linear model u-trees. In: ECML/PKDD (2018)
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2012
Cited alongside, same era.
Alur, R., Bodik, R., Juniwal, G., Martin, M.M.K., Raghothaman, M., Seshia, S.A., Singh, R., Solar-Lezama, A., Torlak, E., Udupa, A.: Syntax-guided synthesis. In: 2013 Formal Methods in Computer-Aided Design. pp. 1–8 (2013), https://ieeexplore.ieee.org/document/6679385
2013
Cited alongside, same era.
Schkufza, E., Sharma, R., Aiken, A.: Stochastic superoptimization. In: Proceedings of the Eighteenth International Conference on Architectural Support for Programming Languages and Operating Systems. p. 305–316. ASPLOS ’13, Association for Computing Machinery, New York, NY, USA (2013), https://doi.org/10.1145/2451116.2451150
2013
Cited alongside, same era.
Udupa, A., Raghavan, A., Deshmukh, J.V., Mador-Haim, S., Martin, M.M., Alur, R.: Transit: Specifying protocols with concolic snippets. SIGPLAN Not. 48
2013
Cited alongside, same era.
2014
Cited alongside, same era.
de Visser, E.J., Cohen, M., Freedy, A., Parasuraman, R.: A design methodology for trust cue calibration in cognitive agents. In: Shumaker, R., Lackey, S. (eds.) Virtual, Augmented and Mixed Reality. Designing and Developing Virtual and Augmented Environments. pp. 251–262. Springer International Publishing, Cham (2014)
2014
Cited alongside, same era.
De Raedt, L., Kersting, K., Natarajan, S., Poole, D.: Statistical Relational Artificial Intelligence: Logic, Probability, and Computation, Synthesis Lectures on Artificial Intelligence and Machine Learning, vol. 32. Morgan & Claypool, San Rafael, CA (2016)
2016
Cited alongside, same era.
Stowers, K., Kasdaglis, N., Newton, O.B., Lakhmani, S.G., Wohleber, R.W., Chen, J.Y.: Intelligent agent transparency. Proceedings of the Human Factors and Ergonomics Society Annual Meeting 60
2016
Cited alongside, same era.
Xu, J., Zhang, Z., Friedman, T., Liang, Y., Broeck, G.: A semantic loss function for deep learning with symbolic knowledge. In: International conference on machine learning. pp. 5502–5511. PMLR (2018)
2018
Later among the works it cites.
Yang, F., Lyu, D., Liu, B., Gustafson, S.: Peorl: Integrating symbolic planning and hierarchical reinforcement learning for robust decision-making. In: Proceedings of the 27th International Joint Conference on Artificial Intelligence. p. 4860–4866. IJCAI’18, AAAI Press, Stockholm, Sweden (2018)
2018
Later among the works it cites.
2018
Later among the works it cites.
Zhang, K., Yang, Z., Liu, H., Zhang, T., Basar, T.: Fully decentralized multi-agent reinforcement learning with networked agents. In: International Conference on Machine Learning. pp. 5872–5881. PMLR (2018)
2018
Later among the works it cites.
Garg, S., Bajpai, A., et al.: Size independent neural transfer for rddl planning. In: Proceedings of the International Conference on Automated Planning and Scheduling. vol. 29, pp. 631–636 (2019)
2019
Later among the works it cites.
Jiang, Z., Luo, S.: Neural logic reinforcement learning. In: Proceedings of the 36th International Conference on Machine Learning. Proceedings of Machine Learning Research, vol. 97, pp. 3110–3119. PMLR, Long Beach, USA (09–15 Jun 2019), https://proceedings.mlr.press/v97/jiang19a.html
2019
Later among the works it cites.
Lyu, D., Yang, F., Liu, B., Gustafson, S.: Sdrl: Interpretable and data-efficient deep reinforcement learning leveraging symbolic planning. In: Proceedings of the Thirty-Third AAAI Conference on Artificial Intelligence and Thirty-First Innovative Applications of Artificial Intelligence Conference and Ninth AAAI Symposium on Educational Advances in Artificial Intelligence. AAAI’19/IAAI’19/EAAI’19, AAAI Press, Honolulu, Hawaii, USA (2019), https://doi.org/10.1609/aaai.v33i01.33012970
2019
Later among the works it cites.
Marra, G., Giannini, F., Diligenti, M., Maggini, M., Gori, M.: T-norms driven loss functions for machine learning. arXiv: Artificial Intelligence (2019)
2019
Later among the works it cites.
Garg, S., Bajpai, A., et al.: Symbolic network: generalized neural policies for relational mdps. In: International Conference on Machine Learning. pp. 3397–3407. PMLR (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
Lamb, L.C., Garcez, A.d., Gori, M., Prates, M.O., Avelar, P.H., Vardi, M.Y.: Graph neural networks meet neural-symbolic computing: A survey and perspective. In: Bessiere, C. (ed.) Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence, IJCAI-20. pp. 4877–4884. International Joint Conferences on Artificial Intelligence Organization, Yokohama, Japan (7 2020), https://doi.org/10.24963/ijcai.2020/679 , survey track
2020
Later among the works it cites.
2020
Later among the works it cites.
Silva, A., Gombolay, M., Killian, T., Jimenez, I., Son, S.H.: Optimization methods for interpretable differentiable decision trees applied to reinforcement learning. In: Chiappa, S., Calandra, R. (eds.) Proceedings of the Twenty Third International Conference on Artificial Intelligence and Statistics. Proceedings of Machine Learning Research, vol. 108, pp. 1855–1865. PMLR, Palermo, Italy (26–28 Aug 2020), https://proceedings.mlr.press/v108/silva20a.html
2020
Later among the works it cites.
Kokel, H., Manoharan, A., Natarajan, S., Ravindran, B., Tadepalli, P.: Reprel: Integrating relational planning and reinforcement learning for effective abstraction. In: Proceedings of the International Conference on Automated Planning and Scheduling. vol. 31, pp. 533–541 (2021)
2021
Later among the works it cites.