Fetching the paper…
Reading the bibliography…
In this article, we study the problem of air-to-ground ultra-reliable and low-latency communication (URLLC) for a moving ground user.
M. Tan, “Multi-agent reinforcement learning: Independent vs. cooperative agents,” in Proceedings of the International Conference on Machine Learning (ICML) , MA, USA, June 1993, pp. 330–337
1993
Earlier work this paper cites.
V. R. Konda and J. N. Tsitsiklis, “Actor-critic algorithms,” in Proceedings of the Advances in Neural Information Processing Systems (NIPS) , CO, USA, January 2000, pp. 1008–1014
2000
Earlier work this paper cites.
2013
Earlier work this paper cites.
S. Sukhbaatar, A. Szlam, and R. Fergus, “Learning multiagent communication with backpropagation,” in Proceedings of the Advances in Neural Information Processing Systems (NIPS) , NY, USA, December 2016, pp. 2252–2260
2016
Earlier work this paper cites.
J. Foerster, I. A. Assael, N. de Freitas, and S. Whiteson, “Learning to communicate with deep multi-agent reinforcement learning,” in Proceedings of the Advances in Neural Information Processing Systems (NIPS) , vol. 29, NY, USA, December 2016, pp. 2145–2153
2016
Earlier work this paper cites.
F. A. Oliehoek and C. Amato, A Concise Introduction to Decentralized POMDPs . Springer Publishing Company, Incorporated, 2016
2016
Earlier work this paper cites.
R. Lowe, Y. Wu, A. Tamar, J. Harb, P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in Proceedings of the Advances in Neural Information Processing Systems (NIPS) , NY, USA, December 2017, pp. 6382–6393
2017
Earlier work this paper cites.
A. Tampuu, T. Matiisen, D. Kodelja, I. Kuzovkin, K. Korjus, J. Aru, J. Aru, and R. Vicente, “Multiagent cooperation and competition with deep reinforcement learning,” PLOS ONE , vol. 12, no. 4, pp. 1–15, 04 2017
2017
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Proceedings of the Advances in Neural Information Processing Systems (NIPS) , CA, USA, December 2017, pp. 5998–6008
2017
Cited alongside, same era.
R. Dey and F. M. Salem, “Gate-variants of gated recurrent unit (GRU) neural networks,” in Proceedings of International Midwest Symposium on Circuits and Systems (MWSCAS) . IEEE, August 2017, pp. 1597–1600
2017
Cited alongside, same era.
H. Kim, J. Park, M. Bennis, and S.-L. Kim, “Massive UAV-to-ground communication and its stable movement control: A mean-field approach,” in Proc. of IEEE SPAWC , Kalamata, Greece, Jun. 2018
2018
Cited alongside, same era.
J. Park, S. Samarakoon, M. Bennis, and M. Debbah, “Wireless network intelligence at the edge,” Proceedings of the IEEE , vol. 107, no. 11, pp. 2204–2239, October 2019
2019
Later among the works it cites.
2020
Later among the works it cites.
Y. Liu, W. Wang, Y. Hu, J. Hao, X. Chen, and Y. Gao, “Multi-agent game abstraction via graph attention neural network,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 05, NY, USA, February 2020, pp. 7211–7218
2020
Later among the works it cites.
S. Jung, W. J. Yun, M. Shin, J. Kim, and J.-H. Kim, “Orchestrated scheduling and multi-agent deep reinforcement learning for cloud-assisted multi-UAV charging systems,” IEEE Transactions on Vehicular Technology , pp. 1–1, 2021
2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Rashid, M. Samvelyan, C. Schroeder, G. Farquhar, J. Foerster, and S. Whiteson, “QMIX: Monotonic value function factorisation for deep multi-agent reinforcement learning,” in Proceedings of the International Conference on Machine Learning (ICML) , Stockholmsmässan, Sweden, July 2018, pp. 4295–4304
2018
Cited alongside, same era.
Z. Zhu, L. Li, and W. Zhou, “QoS-aware 3D deployment of UAV base stations,” in Proceedings of the IEEE International Conference on Wireless Communications and Signal Processing (WCSP) , October 2018, pp. 1–6
2018
Cited alongside, same era.
H. Ji, S. Park, J. Yeo, Y. Kim, J. Lee, and B. Shim, “Ultra-reliable and low-latency communications in 5G downlink: Physical layer aspects,” IEEE Wireless Communications , vol. 25, no. 3, pp. 124–130, June 2018
2018
Cited alongside, same era.
W. J. Yun, S. Yi, and J. Kim, “Multi-agent deep reinforcement learning using attentive graph neural architectures for real-time strategy games,” submitted to IEEE SMC 2021
2021
Closest in time.
C. She, C. Sun, Z. Gu, Y. Li, C. Yang, H. V. Poor, and B. Vucetic, “A tutorial on ultrareliable and low-latency communications in 6G: Integrating domain knowledge into deep learning,” Proceedings of the IEEE , vol. 109, no. 3, pp. 204–246, Mar 2021
2021
Closest in time.