Fetching the paper…
Reading the bibliography…
Many recent breakthroughs in multi-agent reinforcement learning (MARL) require the use of deep neural networks, which are challenging for human experts to interpret and understand.
Shapley, L.: Stochastic games. PNAS 39
1953
Earlier work this paper cites.
Quinlan, J.: Induction of decision trees. Mach. Learning (1986)
1986
Earlier work this paper cites.
Littman, M.: Markov games as a framework for multi-agent reinforcement learning. In: Mach. Learning (1994)
1994
Earlier work this paper cites.
McCallum, R.: Reinforcement learning with selective perception and hidden state. PhD Thesis, Univ. Rochester, Dept. of Comp. Sci. (1997)
1997
Earlier work this paper cites.
Uther, W., Veloso, M.: The lumberjack algorithm for learning linked decision forests. In: Int. Symp. Abstract., Reformulation, and Approx. (2000)
2000
Earlier work this paper cites.
Pyeatt, L., Howe, A.: Decision tree function approximation in reinforcement learning. In: Int. Symp. on Adaptive Syst.: Evol. Comput. and Prob. Graphical Models (2001)
2001
Earlier work this paper cites.
Tuyls, K., et al.: Reinforcement learning in large state spaces. In: Robot Soccer World Cup (2002)
2002
Earlier work this paper cites.
Pyeatt, L.: Reinforcement learning with decision trees. In: Appl. Informatics (2003)
2003
Earlier work this paper cites.
Abbeel, P., Ng, A.: Apprenticeship learning via inverse reinforcement learning. In: ICML (2004)
2004
Earlier work this paper cites.
Ernst, D., et al.: Tree-based batch mode reinforcement learning. JMLR 6
2005
Earlier work this paper cites.
Buciluǎ, C., et al.: Model compression. In: KDD (2006)
2006
Earlier work this paper cites.
Degris, T., et al.: Learning the structure of factored Markov decision processes in reinforcement learning problems. In: ICML (2006)
2006
Earlier work this paper cites.
Strehl, A., et al.: Efficient structure learning in factored-state mdps. In: AAAI (2007)
2007
Earlier work this paper cites.
Oliehoek, F., et al.: Optimal and approximate q-value functions for decentralized pomdps. JAIR 32
2008
Earlier work this paper cites.
Ross, S., et al.: A reduction of imitation learning and structured prediction to no-regret online learning. In: AISTATS (2011)
2011
Earlier work this paper cites.
Matignon, L., et al.: Independent reinforcement learners in cooperative markov games: a survey regarding coordination problems. Knowledge Eng. Review 27
2012
Earlier work this paper cites.
2015
Cited alongside, same era.
Malialis, K., Kudenko, D.: Distributed response to network intrusions using multiagent reinforcement learning. Eng. Appl. Artif. Intell. (2015)
2015
Cited alongside, same era.
Foerster, J., et al.: Stabilising experience replay for deep multi-agent reinforcement learning. In: ICML (2017)
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Paszke, A., et al.: Automatic differentiation in pytorch (2017)
2017
Cited alongside, same era.
Molnar, C.: Interpretable Machine Learning (2019)
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
Bhalla, S., et al.: Deep multi agent reinforcement learning for autonomous driving. In: Canadian Conf. Artif. Intell. (2020)
2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
Bastani, O., et al.: Verifiable reinforcement learning via policy extraction. In: NeurIPS (2018)
2018
Cited alongside, same era.
Foerster, J., et al.: Counterfactual multi-agent policy gradients. In: AAAI (2018)
2018
Cited alongside, same era.
Lipton, Z.: The mythos of model interpretability. ACM Queue 16
2018
Cited alongside, same era.
Rashid, T., et al.: Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning. In: ICML (2018)
2018
Cited alongside, same era.
Wang, T., et al.: Dataset distillation. arXiv preprint arXiv:1811.10959 (2018)
2018
Cited alongside, same era.
Berner, C., et al.: Dota 2 with large scale deep reinforcement learning. arXiv preprint 1912.06680 (2019)
2019
Cited alongside, same era.
Later among the works it cites.
Kazhdan, D., et al.: Marleme: A multi-agent reinforcement learning model extraction library. In: IJCNN (2020)
2020
Later among the works it cites.
Meng, Z., et al.: Interpreting deep learning-based networking systems. In: Proceedings of the Annual conference of the ACM Special Interest Group on Data Communication on the applications, technologies, architectures, and protocols for computer communication (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
2021
Later among the works it cites.
Motokawa, Y., Sugawara, T.: Mat-dqn: Toward interpretable multi-agent deep reinforcement learning for coordinated activities. In: ICANN (2021)
2021
Later among the works it cites.
Topin, N., et al.: Iterative bounding mdps: Learning interpretable policies via non-interpretable methods. In: AAAI (2021)
2021
Later among the works it cites.
2021
Later among the works it cites.
Heuillet, A., et al.: Collective explainable ai: Explaining cooperative strategies and agent contribution in multiagent reinforcement learning with shapley values. IEEE Comput. Intell. Magazine 17
2022
Closest in time.
2022
Closest in time.