Fetching the paper…
Reading the bibliography…
Recent years have witnessed the great success of multi-agent systems (MAS).
C. Berge, Hypergraphs: combinatorics of finite sets . Elsevier, 1984, vol. 45
1984
Earlier work this paper cites.
R. N. Bracewell and R. N. Bracewell, The Fourier transform and its applications . McGraw-Hill New York, 1986, vol. 31999
1986
Earlier work this paper cites.
2006
Earlier work this paper cites.
J.-M. Lasry and P.-L. Lions, “Mean field games,” Japanese journal of mathematics , vol. 2, no. 1, pp. 229–260, 2007
2007
Earlier work this paper cites.
F. Scarselli, M. Gori, A. C. Tsoi et al. , “The graph neural network model,” IEEE transactions on neural networks , vol. 20, no. 1, pp. 61–80, 2008
2008
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver et al. , “Human-level control through deep reinforcement learning,” nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
M. Hausknecht and P. Stone, “Deep recurrent q-learning for partially observable mdps,” in 2015 aaai fall symposium series , 2015
2015
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison et al. , “Mastering the game of go with deep neural networks and tree search,” nature , vol. 529, no. 7587, pp. 484–489, 2016
2016
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell et al. , “End-to-end training of deep visuomotor policies,” The Journal of Machine Learning Research , vol. 17, no. 1, pp. 1334–1373, 2016
2016
Earlier work this paper cites.
S. Sukhbaatar, R. Fergus et al. , “Learning multiagent communication with backpropagation,” Advances in neural information processing systems , vol. 29, pp. 2244–2252, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
M. Defferrard, X. Bresson, and P. Vandergheynst, “Convolutional neural networks on graphs with fast localized spectral filtering,” Advances in neural information processing systems , vol. 29, pp. 3844–3852, 2016
2016
Earlier work this paper cites.
F. A. Oliehoek and C. Amato, A concise introduction to decentralized POMDPs . Springer, 2016
2016
Earlier work this paper cites.
A. Tampuu, T. Matiisen, D. Kodelja et al. , “Multiagent cooperation and competition with deep reinforcement learning,” PloS one , vol. 12, no. 4, p. e0172395, 2017
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
W. L. Hamilton, R. Ying et al. , “Inductive representation learning on large graphs,” in Proceedings of the 31st International Conference on Neural Information Processing Systems , 2017, pp. 1025–1035
2017
Cited alongside, same era.
2018
Cited alongside, same era.
T. Rashid, M. Samvelyan, C. Schroeder et al. , “Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning,” in International Conference on Machine Learning . PMLR, 2018, pp. 4295–4304
2018
Cited alongside, same era.
K. Son, D. Kim, W. J. Kang et al. , “Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning,” in International Conference on Machine Learning . PMLR, 2019, pp. 5887–5896
2019
Later among the works it cites.
J. Jiang, Y. Wei, Y. Feng et al. , “Dynamic hypergraph neural networks.” in IJCAI , 2019, pp. 2635–2641
2019
Later among the works it cites.
C. Gong, Y. Bai, X. Hou, and X. Ji, “Stable training of bellman error in reinforcement learning,” in International Conference on Neural Information Processing . Springer, 2020, pp. 439–448
2020
Later among the works it cites.
2020
Later among the works it cites.
T. Tang Nguyen, S. Gupta, and S. Venkatesh, “Distributional reinforcement learning with maximum mean discrepancy,” arXiv e-prints , pp. arXiv–2007, 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Y. Yang, R. Luo, M. Li et al. , “Mean field multi-agent reinforcement learning,” in International Conference on Machine Learning . PMLR, 2018, pp. 5571–5580
2018
Cited alongside, same era.
T. H. Nguyen and R. Grishman, “Graph convolutional networks with argument-aware pooling for event detection,” in Thirty-second AAAI conference on artificial intelligence , 2018
2018
Cited alongside, same era.
Z. Wang, Q. Lv, X. Lan et al. , “Cross-lingual knowledge graph alignment via graph convolutional networks,” in Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing , 2018, pp. 349–357
2018
Cited alongside, same era.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Cited alongside, same era.
M. Rahmati, M. Nadeem, V. Sadhu et al. , “Uw-marl: Multi-agent reinforcement learning for underwater adaptive sampling using autonomous vehicles,” in Proceedings of the International Conference on Underwater Networks & Systems , 2019, pp. 1–5
2019
Cited alongside, same era.
J. Cui, Y. Liu, and A. Nallanathan, “Multi-agent reinforcement learning-based resource allocation for uav networks,” IEEE Transactions on Wireless Communications , vol. 19, no. 2, pp. 729–743, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
C. Gong, Q. He, Y. Bai, X. Hou, G. Fan, and Y. Liu, “Wide-sense stationary policy optimization with bellman residual on video games,” in 2021 IEEE International Conference on Multimedia and Expo, ICME 2021, Shenzhen, China, July 5-9, 2021 . IEEE, 2021, pp. 1–6
2021
Closest in time.
T. Zhang, Y. Li, C. Wang et al. , “Fop: Factorizing optimal joint policy of maximum-entropy multi-agent reinforcement learning,” in International Conference on Machine Learning . PMLR, 2021, pp. 12 491–12 500
2021
Closest in time.
S. Bai, F. Zhang, and P. H. Torr, “Hypergraph convolution and hypergraph attention,” Pattern Recognition , vol. 110, p. 107637, 2021
2021
Closest in time.
2021
Closest in time.