Fetching the paper…
Reading the bibliography…
Many scenarios in mobility and traffic involve multiple different agents that need to cooperate to find a joint solution.
L. Shapley, “A value for n-person games,” Ann. Math. Study28, Contributions to the Theory of Games , pp. 307–317, 1953
1953
Earlier work this paper cites.
M. Tan, “Multi-agent reinforcement learning: Independent vs. cooperative agents,” Int. Conf. on Machine Learning , pp. 330–337, 1993
1993
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning - an introduction , ser. Adaptive computation and machine learning. MIT Press, 1998
1998
Earlier work this paper cites.
T. Balch, “Reward and diversity in multirobot foraging,” Work. Agents Learning About, From and With other Agents , pp. 92–99, 1999
1999
Earlier work this paper cites.
M. L. Littman, “Value-function reinforcement learning in Markov games,” Cognitive Systems Research , vol. 2, no. 1, pp. 55–66, 2001
2001
Earlier work this paper cites.
D. H. Wolpert and K. Tumer, “Optimal payoff functions for members of collectives,” Adv. in Compl. Sys. , vol. 4, no. 2/3, pp. 265–279, 2001
2001
Earlier work this paper cites.
D. W. Casbeer, D. B. Kingston, R. W. Beard, and T. W. McLain, “Cooperative forest fire surveillance using a team of small unmanned air vehicles,” Int. J. Syst. Sci. , vol. 37, no. 6, pp. 351–360, 2006
2006
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
J. N. Foerster, Y. M. Assael, N. De Freitas, and S. Whiteson, “Learning to communicate with deep multi-agent reinforcement learning,” in NeurIPS , 2016, pp. 2145–2153
2016
Earlier work this paper cites.
S. Sukhbaatar, A. Szlam, and R. Fergus, “Learning multiagent communication with backpropagation,” in Advances in Neural Information Processing Systems , 2016, pp. 2252–2260
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
R. Lowe, Y. Wu, A. Tamar, J. Harb, P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in NeurIPS , 2017, pp. 6380–6391
2017
Earlier work this paper cites.
J. Hao, D. Huang, Y. Cai, and H. fung Leung, “The dynamics of reinforcement social learning in networked cooperative multiagent systems,” Eng. Applicat. of Artificial Intellig. , vol. 58, pp. 111–122, 2017
2017
Earlier work this paper cites.
S. Shah, D. Dey, C. Lovett, and A. Kapoor, “Airsim: High-fidelity visual and physical simulation for autonomous vehicles,” in Intl. Conf. Field and Serv. Robot. , Zurich, Switzerland, 2017, pp. 621–635
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
P. A. Lopez, M. Behrisch, L. Bieker-Walz, J. Erdmann, Y.-P. Flötteröd, R. Hilbrich, L. Lücken, J. Rummel, P. Wagner, and E. Wießner, “Microscopic traffic simulation using sumo,” in The 21st IEEE International Conference on Intelligent Transportation Systems . IEEE, 2018
2018
Earlier work this paper cites.
J. N. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson, “Counterfactual multi-agent policy gradients,” in AAAI Conf. on Artificial Intelligence , 2018, pp. 2974–2982
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
J. Jiang and Z. Lu, “Learning attentional communication for multi-agent cooperation,” in NeurIPS , 2018, pp. 7254–7264
2018
Earlier work this paper cites.
Y. Yang, R. Luo, M. Li, M. Zhou, W. Zhang, and J. Wang, “Mean field multi-agent reinforcement learning,” in Int. Conf. on Machine Learning , vol. 12, 2018, pp. 5571–5580
2018
Earlier work this paper cites.
E. Liang, R. Liaw, R. Nishihara, P. Moritz, R. Fox, K. Goldberg, J. Gonzalez, M. Jordan, and I. Stoica, “RLlib: Abstractions for distributed reinforcement learning,” in Int. Conf. on Machine Learning , 2018, pp. 3053–3062
2018
Earlier work this paper cites.
L. Zheng, J. Yang, H. Cai, W. Zhang, J. Wang, and Y. Yu, “MAgent: a many-agent reinforcement learning platform for artificial collective intelligence,” in AAAI Conf. Artificial Intellig. , 2018, pp. 8222–8223
2018
Earlier work this paper cites.
E. Leurent, “An environment for autonomous driving decision-making,” https://github.com/eleurent/highway-env , 2018
2018
Cited alongside, same era.
M. Elloumi, R. Dhaou, B. Escrig, H. Idoudi, and L. A. Saïdane, “Monitoring road traffic with a uav-based system,” in IEEE Wireless Communications and Networking Conf. , 2018
2018
Cited alongside, same era.
Y. Yang, Z. Zheng, K. Bian, L. Song, and Z. Han, “Real-time profiling of fine-grained air quality index distribution using UAV sensing,” IEEE Internet Things J. , vol. 5, no. 1, pp. 186–198, 2018
2018
Cited alongside, same era.
A. Oroojlooyjadid and D. Hajinezhad, “A review of cooperative multi-agent deep reinforcement learning,” arXiv preprint 1908.03963 , 2019
2019
Cited alongside, same era.
H. Zhang, S. Feng, C. Liu, Y. Ding, Y. Zhu, Z. Zhou, W. Zhang, Y. Yu, H. Jin, and Z. Li, “Cityflow: A multi-agent reinforcement learning environment for large scale city traffic scenario,” in World Wide Web Conf. , San Francisco, CA, 2019, pp. 3620–3624
J. Bernhard, K. Esterle, P. Hart, and T. Kessler, “BARK: open behavior benchmarking in multi-agent environments,” in IEEE/RSJ Int. Conf. Intelligent Robots and Systems , Las Vegas, NV, 2020, pp. 6201–6208
2020
Later among the works it cites.
P. Palanisamy, “Multi-agent connected autonomous driving using deep reinforcement learning,” in Int. Joint Conf. on Neural Networks , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
G. Kontes, D. Scherer, T. Nisslbeck, J. Fischer, and C. Mutschler, “High-speed collision avoidance using deep reinforcement learning and domain randomization for autonomous vehicles,” in IEEE Int. Conf. Intell. Transportation Systems , 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
P. Hernandez-Leal, B. Kartal, and M. E. Taylor, “A survey and critique of multiagent deep reinforcement learning,” Autonomous Agents and Multi-Agent Systems , vol. 33, no. 6, pp. 750–797, nov 2019
2019
Cited alongside, same era.
A. Mahajan, T. Rashid, M. Samvelyan, and S. Whiteson, “MAVEN: multi-agent variational exploration,” in NeurIPS , 2019
2019
Cited alongside, same era.
A. Singh, T. Jain, and S. Sukhbaatar, “Learning when to communicate at scale in multiagent cooperative and competitive tasks,” in Int. Conf. on Learning Representations , 2019
2019
Cited alongside, same era.
X. Li, J. Zhang, J. Bian, Y. Tong, and T.-Y. Liu, “A cooperative multi-agent reinforcement learning framework for resource balancing in complex logistics network,” in Int. Joint Conf. on Autonomous Agents and Multiagent Systems , Mar. 2019
2019
Cited alongside, same era.
Y. Wu, H. Chen, and F. Zhu, “Dcl-aim: Decentralized coordination learning of autonomous intersection management for connected and automated vehicles,” Transportation Research Part C: Emerging Technologies , vol. 103, pp. 246–260, 2019
2019
Cited alongside, same era.
J. Cui, Y. Liu, and A. Nallanathan, “The application of multi-agent reinforcement learning in UAV networks,” in IEEE Int. Conf. on Communications Workshops . Shanghai, China: IEEE, May 2019
2019
Cited alongside, same era.
M. Mozaffari, W. Saad, M. Bennis, Y. Nam, and M. Debbah, “A tutorial on uavs for wireless networks: Applications, challenges, and open problems,” IEEE Commun. Surv. Tutorials , vol. 21, no. 3, pp. 2334–2360, 2019
2019
Cited alongside, same era.
B. Osinski, A. Jakubowski, P. Ziecina, P. Milos, S. Galias, C. Homoceanu, and H. Michalewski, “Simulation-based reinforcement learning for real-world autonomous driving,” in IEEE Int. Conf. Robotics and Automation , May 2020, pp. 6411–6418
2020
Later among the works it cites.
K. Zhang, Z. Yang, and T. Başar, “Multi-agent reinforcement learning: A selective overview of theories and algorithms,” in Studies in Systems, Decision and Control , 2021, vol. 325, pp. 321–384
2021
Later among the works it cites.
L. Canese, G. C. Cardarilli, L. Di Nunzio, R. Fazzolari, D. Giardino, M. Re, and S. Spanò, “Multi-agent reinforcement learning: A review of challenges and applications,” Appl. Sciences , vol. 11, no. 11, 2021
2021
Later among the works it cites.
J. Li, K. Kuang, B. Wang, F. Liu, L. Chen, F. Wu, and J. Xiao, “Shapley counterfactual credits for multi-agent reinforcement learning,” in ACM Intl. Conf. Knowl. Disc. and Data Min. , 2021, pp. 934–942
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
E. Kargar and V. Kyrki, “MACRPO: multi-agent cooperative recurrent policy optimization,” arXiv preprint 2109.00882 , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
J. Panerati, H. Zheng, S. Zhou, J. Xu, A. Prorok, and A. P. Schoellig, “Learning to fly—a gym environment with pybullet physics for reinforcement learning of multi-agent quadcopter control,” in IEEE/RSJ Int. Conf. on Intelligent Robots and Systems (IROS) , 2021
2021
Later among the works it cites.
Z. Wang, H. Zhu, M. He, Y. Zhou, X. Luo, and N. Zhang, “Gan and multi-agent DRL based decentralized traffic light signal control,” IEEE Trans. Veh. Technol. , 2021
2021
Later among the works it cites.
M. Wang, L. Wu, J. Li, and L. He, “Traffic signal control with reinforcement learning based on region-aware cooperative strategy,” IEEE Transactions on Intelligent Transportation Systems , 2021
2021
Later among the works it cites.
L. M. Schmidt, G. Kontes, A. Plinge, and C. Mutschler, “Can you trust your autonomous car? interpretable and verifiably safe reinforcement learning,” in IEEE Intelligent Vehicles Symp. , Nagoya, Japan, Jul. 2021, pp. 171–178
2021
Later among the works it cites.
S. Jung, W. J. Yun, J. Kim, J.-H. Kim, and F. Falcone, “Coordinated multi-agent deep reinforcement learning for energy-aware UAV-based big-data platforms,” mdpi.com , 2021
2021
Later among the works it cites.
C. Liu, C. X. Chen, and C. Chen, “META: a city-wide taxi repositioning framework based on multi-agent reinforcement learning,” IEEE Trans. on Intelligent Transportation Systems , 2021
2021
Later among the works it cites.
2022
Closest in time.
M. Stahlke, T. Feigl, M. H. Castaneda Garcia, R. S. Gallacher, J. Seitz, and C. Mutschler, “Transfer learning to adapt 5G fingerprint-based localization across environments,” in IEEE Veh. Tech. Conf. , 2022
2022
Closest in time.
P. Sunehag, G. Lever, A. Gruslys, W. M. Czarnecki, V. Zambaldi, M. Jaderberg, M. Lanctot, N. Sonnerat, J. Z. Leibo, K. Tuyls, and T. Graepel, “Value-decomposition networks for cooperative multi-agent learning based on team reward,” in Intl. Jnt. Conf. Autonomous Agents and Multiagent Systems , 2018, pp. 2085–2087
2087
Closest in time.