Fetching the paper…
Reading the bibliography…
Trading off performance guarantees in favor of scalability, the Multi-Agent Path Finding (MAPF) community has recently started to embrace Multi-Agent Reinforcement Learning (MARL), where agents learn to collaboratively generate individual, collision-free (but often suboptimal) paths.
M. Erdmann and T. Lozano-Perez, “On multiple moving objects,” Algorithmica , vol. 2, pp. 477–521, 1987
1987
Earlier work this paper cites.
Q. Sajid, R. Luna, and K. Bekris, “Multi-agent pathfinding with simultaneous execution of single-agent primitives,” in International symposium on combinatorial search , vol. 3, no. 1, 2012
2012
Earlier work this paper cites.
G. Wagner and H. Choset, “Subdimensional expansion for multirobot path planning,” Artificial intelligence , vol. 219, pp. 1–24, 2015
2015
Earlier work this paper cites.
G. Sharon, R. Stern, A. Felner, and N. R. Sturtevant, “Conflict-based search for optimal multi-agent pathfinding,” Artificial Intelligence , vol. 219, pp. 40–66, 2015
2015
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson, “Counterfactual multi-agent policy gradients,” in Proceedings of the AAAI conference on artificial intelligence , vol. 32, no. 1, 2018
2018
Earlier work this paper cites.
T. Rashid, M. Samvelyan, C. Schroeder, G. Farquhar, J. Foerster, and S. Whiteson, “Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning,” in ICML , 2018, pp. 4295–4304
2018
Earlier work this paper cites.
2018
Cited alongside, same era.
G. Sartoretti, J. Kerr, Y. Shi, G. Wagner, T. S. Kumar, S. Koenig, and H. Choset, “PRIMAL: Pathfinding via Reinforcement and Imitation Multi-Agent Learning,” IEEE Robotics and Automation Letters (RA-L) , vol. 4(3), pp. 2378–2385, 2019
2019
Cited alongside, same era.
S. Q. Zhang, Q. Zhang, and J. Lin, “Efficient communication in multi-agent reinforcement learning via variance based control,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Cited alongside, same era.
R. Stern, N. Sturtevant, A. Felner, S. Koenig, H. Ma, T. Walker, J. Li, D. Atzmon, L. Cohen, T. Kumar, et al. , “Multi-agent pathfinding: Definitions, variants, and benchmarks,” in Proceedings of SoCS , vol. 10, no. 1, 2019, pp. 151–158
2019
Cited alongside, same era.
Z. Ma, Y. Luo, and H. Ma, “Distributed heuristic multi-agent path finding with communication,” in ICRA , 2021, pp. 8699–8705
2021
Later among the works it cites.
Z. Ma, Y. Luo, and J. Pan, “Learning selective communication for multi-agent path finding,” IEEE RA-L , vol. 7(2), pp. 1455–1462, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
Y. Wang, M. Damani, P. Wang, Y. Cao, and G. Sartoretti, “Distributed reinforcement learning for robot teams: A review,” Conditionally Accepted to Springer’s Current Robotics Reports , 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Parisotto, F. Song, J. Rae, R. Pascanu, C. Gulcehre, S. Jayakumar, M. Jaderberg, R. L. Kaufman, A. Clark, S. Noury, et al. , “Stabilizing transformers for reinforcement learning,” in International conference on machine learning . PMLR, 2020, pp. 7487–7498
2020
Cited alongside, same era.
J. Li, G. Gange, D. Harabor, P. J. Stuckey, H. Ma, and S. Koenig, “New techniques for pairwise symmetry breaking in multi-agent path finding,” in ICAPS , vol. 30, 2020, pp. 193–201
2020
Cited alongside, same era.
A. P. Badia, B. Piot, S. Kapturowski, P. Sprechmann, A. Vitvitskyi, Z. D. Guo, and C. Blundell, “Agent57: Outperforming the atari human benchmark,” in International Conference on Machine Learning . PMLR, 2020, pp. 507–517
2020
Cited alongside, same era.
M. Damani, Z. Luo, E. Wenzel, and G. Sartoretti, “PRIMAL 2 : Pathfinding via Reinforcement and Imitation Multi-Agent Learning - Lifelong,” IEEE RA-L , vol. 6, no. 2, pp. 2666–2673, 2021
2021
Cited alongside, same era.
Q. Li, W. Lin, Z. Liu, and A. Prorok, “Message-aware graph attention networks for large-scale multi-robot path planning,” IEEE Robotics and Automation Letters , vol. 6, no. 3, pp. 5533–5540, 2021
2021
Cited alongside, same era.
2022
Later among the works it cites.
S. Shaw, E. Wenzel, A. Walker, and G. Sartoretti, “ForMIC: Foraging via Multiagent rl with Implicit Communication,” IEEE Robotics and Automation Letters , vol. 7, no. 2, pp. 4877–4884, 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Wang and G. Sartoretti, “FCMNet: Full Communication Memory Net for team-level cooperation in multi-agent systems,” in AAMAS , 2022, pp. 1355–1363
2022
Later among the works it cites.
Y. Wang, B. Xiang, S. Huan, and G. Sartoretti, “SCRIMP: Scalable communication for reinforcement- and imitation-learning-based multi-agent pathfinding,” in Accepted as an extended abstract
2023
Closest in time.