Fetching the paper…
Reading the bibliography…
We consider a fully cooperative multi-agent system where agents cooperate to maximize a system's utility in a partial-observable environment.
Williams, R.J.: Simple statistical gradient-following algorithms for connectionist reinforcement learning. Machine learning 8
1992
Earlier work this paper cites.
Tan, M.: Multi-agent reinforcement learning: Independent vs. cooperative agents. In: Proceedings of the tenth international conference on machine learning. pp. 330–337 (1993)
1993
Earlier work this paper cites.
Wolpert, D.H., Tumer, K.: Optimal payoff functions for members of collectives. In: Modeling complexity in economic and social systems, pp. 355–369. World Scientific (2002)
2002
Earlier work this paper cites.
Bu, L., Babu, R., De Schutter, B., et al.: A comprehensive survey of multiagent reinforcement learning. IEEE Transactions on Systems, Man, and Cybernetics, Part C (Applications and Reviews) 38
2008
Earlier work this paper cites.
Bengio, Y., Louradour, J., Collobert, R., Weston, J.: Curriculum learning. In: Proceedings of the 26th annual international conference on machine learning. pp. 41–48 (2009)
2009
Earlier work this paper cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., Hassabis, D.: Human-level control through deep reinforcement learning. Nature 518
2015
Earlier work this paper cites.
Choo, B.Y., Adams, S.C., Weiss, B.A., Marvel, J.A., Beling, P.A.: Adaptive multi-scale prognostics and health management for smart manufacturing systems. International journal of prognostics and health management 7
2016
Earlier work this paper cites.
Foerster, J., Assael, I.A., De Freitas, N., Whiteson, S.: Learning to communicate with deep multi-agent reinforcement learning. In: Advances in neural information processing systems. pp. 2137–2145 (2016)
2016
Earlier work this paper cites.
2016
Cited alongside, same era.
Mnih, V., Badia, A.P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., Kavukcuoglu, K.: Asynchronous methods for deep reinforcement learning. In: International conference on machine learning. pp. 1928–1937 (2016)
2016
Cited alongside, same era.
Sukhbaatar, S., Fergus, R., et al.: Learning multiagent communication with backpropagation. In: Advances in neural information processing systems. pp. 2244–2252 (2016)
2016
Cited alongside, same era.
Choo, B.Y., Adams, S., Beling, P.: Health-aware hierarchical control for smart manufacturing using reinforcement learning. In: 2017 IEEE International Conference on Prognostics and Health Management (ICPHM). pp. 40–47. IEEE (2017)
2017
Cited alongside, same era.
2017
Later among the works it cites.
Foerster, J.N., Farquhar, G., Afouras, T., Nardelli, N., Whiteson, S.: Counterfactual multi-agent policy gradients. In: Thirty-second AAAI conference on artificial intelligence (2018)
2018
Later among the works it cites.
2018
Later among the works it cites.
Jiang, J., Lu, Z.: Learning attentional communication for multi-agent cooperation. In: Advances in neural information processing systems. pp. 7254–7264 (2018)
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gupta, J.K., Egorov, M., Kochenderfer, M.: Cooperative multi-agent control using deep reinforcement learning. In: International Conference on Autonomous Agents and Multiagent Systems. pp. 66–83. Springer (2017)
2017
Cited alongside, same era.
Hamilton, W., Ying, Z., Leskovec, J.: Inductive representation learning on large graphs. In: Advances in neural information processing systems. pp. 1024–1034 (2017)
2017
Cited alongside, same era.
Lowe, R., Wu, Y., Tamar, A., Harb, J., Abbeel, O.P., Mordatch, I.: Multi-agent actor-critic for mixed cooperative-competitive environments. In: Advances in neural information processing systems. pp. 6379–6390 (2017)
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Later among the works it cites.
Das, A., Gervet, T., Romoff, J., Batra, D., Parikh, D., Rabbat, M., Pineau, J.: Tarmac: Targeted multi-agent communication. In: International Conference on Machine Learning. pp. 1538–1546 (2019)
2019
Later among the works it cites.
Samvelyan, M., Rashid, T., Schroeder de Witt, C., Farquhar, G., Nardelli, N., Rudner, T.G., Hung, C.M., Torr, P.H., Foerster, J., Whiteson, S.: The starcraft multi-agent challenge. In: Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems. pp. 2186–2188. International Foundation for Autonomous Agents and Multiagent Systems (2019)
2019
Later among the works it cites.