Fetching the paper…
Reading the bibliography…
In recent years, reinforcement learning and its multi-agent analogue have achieved great success in solving various complex control problems.
1910
Earlier work this paper cites.
1912
Earlier work this paper cites.
K. Cui and H. Koeppl, “Approximately solving mean field games via entropy-regularized deep reinforcement learning,” in Proc. AISTATS , 2021, pp. 1909–1917
1917
Earlier work this paper cites.
O. Khatib, “Real-time obstacle avoidance for manipulators and mobile robots,” in Proc. IEEE ICRA , vol. 2, 1985, pp. 500–505
1985
Earlier work this paper cites.
R. A. DeVore and G. G. Lorentz, Constructive approximation . Springer Science & Business Media, 1993, vol. 303
1993
Earlier work this paper cites.
M. Tan, “Multi-agent reinforcement learning: Independent vs. cooperative agents,” in Proc. ICML , 1993, pp. 330–337
1993
Earlier work this paper cites.
P. Fiorini and Z. Shiller, “Motion planning in dynamic environments using velocity obstacles,” Int. J. Robot. Res. , vol. 17, no. 7, pp. 760–772, 1998
1998
Earlier work this paper cites.
K. Lerman, A. Galstyan, A. Martinoli, and A. Ijspeert, “A macroscopic analytical model of collaboration in distributed robotic systems,” Artif. Life , vol. 7, no. 4, pp. 375–393, 2001
2001
Earlier work this paper cites.
P. Vincent and I. Rubin, “A framework and analysis for cooperative search using UAV swarms,” in Proc. ACM Symp. Appl. Comput. , 2004, pp. 79–86
2004
Earlier work this paper cites.
K. R. Parthasarathy, Probability measures on metric spaces . American Mathematical Soc., 2005, vol. 352
2005
Earlier work this paper cites.
N. Correll and A. Martinoli, “System identification of self-organizing robotic swarms,” in Distributed Autonomous Robotic Systems 7 . Springer, 2006, pp. 31–40
2006
Earlier work this paper cites.
D. Milutinović and P. Lima, “Modeling and optimal centralized control of a large-size robotic population,” IEEE Trans. Robot. , vol. 22, no. 6, pp. 1280–1285, 2006
2006
Earlier work this paper cites.
M. Huang, R. P. Malhamé, P. E. Caines et al. , “Large population stochastic dynamic games: closed-loop mckean-vlasov systems and the nash certainty equivalence principle,” Commun. Inf. Syst. , vol. 6, no. 3, pp. 221–252, 2006
2006
Earlier work this paper cites.
J.-M. Lasry and P.-L. Lions, “Mean field games,” Japanese J. Math. , vol. 2, no. 1, pp. 229–260, 2007
2007
Earlier work this paper cites.
R. Gross and M. Dorigo, “Evolution of solitary and group transport behaviors for autonomous robots capable of self-assembling,” Adapt. Behav. , vol. 16, no. 5, pp. 285–305, 2008
2008
Earlier work this paper cites.
H. Hamann and H. Wörn, “A framework of space–time continuous models for algorithm design in swarm robotics,” Swarm Intell. , vol. 2, no. 2, pp. 209–239, 2008
2008
Earlier work this paper cites.
——, “Towards group transport by swarms of robots,” Int. J. Bio-Inspired Comput. , vol. 1, no. 1/2, pp. 1–13, 2009
2009
Earlier work this paper cites.
C. Villani, Optimal transport: old and new . Springer, 2009, vol. 338
2009
Earlier work this paper cites.
2011
Earlier work this paper cites.
M. Brambilla, E. Ferrante, M. Birattari, and M. Dorigo, “Swarm robotics: a review from the swarm engineering perspective,” Swarm Intell. , vol. 7, no. 1, pp. 1–41, 2013
2013
Earlier work this paper cites.
A. Bensoussan, J. Frehse, P. Yam et al. , Mean field games and mean field type control theory . Springer, 2013, vol. 101
2013
Earlier work this paper cites.
H. A. Al-Rawi, M. A. Ng, and K.-L. A. Yau, “Application of reinforcement learning to routing in distributed wireless networks: A review,” Artif. Intell. Rev. , vol. 43, no. 3, pp. 381–416, 2015
2015
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in Proc. ICML . PMLR, 2015, pp. 1889–1897
2015
Earlier work this paper cites.
K. Elamvazhuthi, M. Kawski, S. Biswal, V. Deshmukh, and S. Berman, “Mean-field controllability and decentralized stabilization of markov chains,” in Proc. IEEE CDC , 2017, pp. 3131–3137
2017
Earlier work this paper cites.
U. Eren and B. Açıkmeşe, “Velocity field generation for density control of swarms using heat equation and smoothing kernels,” IFAC-PapersOnLine , vol. 50, no. 1, pp. 9405–9411, 2017
2017
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Process. Mag. , vol. 34, no. 6, pp. 26–38, 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
J. K. Gupta, M. Egorov, and M. Kochenderfer, “Cooperative multi-agent control using deep reinforcement learning,” in Proc. AAMAS , 2017, pp. 66–83
2017
Cited alongside, same era.
W. Giernacki, M. Skwierczyński, W. Witwicki, P. Wroński, and P. Kozierski, “Crazyflie 2.0 quadrotor as a platform for research and education in robotics and control engineering,” in Proc. IEEE MMAR Conf. , 2017, pp. 37–42
2017
Cited alongside, same era.
P. E. Caines and M. Huang, “Graphon mean field games and the GMFG equations: ε \varepsilon -Nash equilibria,” in Proc. IEEE CDC , 2019, pp. 286–292
2019
Later among the works it cites.
M. Schranz, M. Umlauft, M. Sende, and W. Elmenreich, “Swarm robotic behaviors and current applications,” Front. Robot. AI , vol. 7, p. 36, 2020
2020
Later among the works it cites.
M. Dorigo, G. Theraulaz, and V. Trianni, “Reflections on the future of swarm robotics,” Science Robotics , vol. 5, no. 49, p. eabe4385, 2020
2020
Later among the works it cites.
K. Zhang, Z. Yang, and T. Başar, “Multi-agent reinforcement learning: A selective overview of theories and algorithms,” Handbook of Reinforcement Learning and Control , pp. 321–384, 2021
2021
Later among the works it cites.
T. Zheng, Q. Han, and H. Lin, “Transporting robotic swarms via mean-field feedback control,” IEEE Trans. Autom. Control , 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S.-J. Chung, A. A. Paranjape, P. Dames, S. Shen, and V. Kumar, “A survey on aerial swarm robotics,” IEEE Trans. Robot. , vol. 34, no. 4, pp. 837–855, 2018
2018
Cited alongside, same era.
E. Tuci, M. H. Alkilabi, and O. Akanyeti, “Cooperative object transport in multi-robot systems: A review of the state-of-the-art,” Front. Robot. AI , vol. 5, p. 59, 2018
2018
Cited alongside, same era.
D. Albani, T. Manoni, D. Nardi, and V. Trianni, “Dynamic UAV swarm deployment for non-uniform coverage,” in Proc. AAMAS , 2018, pp. 523–531
2018
Cited alongside, same era.
H. Hamann, Swarm Robotics: A Formal Approach . Springer, 2018
2018
Cited alongside, same era.
V. Deshmukh, K. Elamvazhuthi, S. Biswal, Z. Kakish, and S. Berman, “Mean-field stabilization of markov chain models for robotic swarms: Computational approaches and experimental results,” IEEE Robot. Autom. Lett. , vol. 3, no. 3, pp. 1985–1992, 2018
2018
Cited alongside, same era.
K. Elamvazhuthi, S. Biswal, and S. Berman, “Mean-field stabilization of robotic swarms to probability distributions with disconnected supports,” in Proc. IEEE ACC , 2018, pp. 885–892
2018
Cited alongside, same era.
S. Mayya, P. Pierpaoli, G. Nair, and M. Egerstedt, “Localization in densely packed swarms using interrobot collisions as a sensing modality,” IEEE Trans. Robot. , vol. 35, no. 1, pp. 21–34, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2021
Later among the works it cites.
H. Gu, X. Guo, X. Wei, and R. Xu, “Mean-field controls with Q-learning for cooperative MARL: convergence and complexity analysis,” SIAM J. Math. Data Sci. , vol. 3, no. 4, pp. 1168–1196, 2021
2021
Later among the works it cites.
M. Everett, Y. F. Chen, and J. P. How, “Collision avoidance in pedestrian-rich environments with deep reinforcement learning,” IEEE Access , vol. 9, pp. 10 357–10 377, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
G. Papoudakis, F. Christianos, L. Schäfer, and S. V. Albrecht, “Benchmarking multi-agent deep reinforcement learning algorithms in cooperative tasks,” in Proc. NeurIPS Datasets and Benchmarks , 2021
2021
Later among the works it cites.
J. Pérolat, S. Perrin, R. Elie, M. Laurière, G. Piliouras, M. Geist, K. Tuyls, and O. Pietquin, “Scaling mean field games by online mirror descent,” in Proc. AAMAS , vol. 21, 2022, pp. 1028–1037
2022
Closest in time.
2022
Closest in time.
B. Anahtarci, C. D. Kariksiz, and N. Saldi, “Q-learning in regularized mean-field games,” Dyn. Games and Appl. , pp. 1–29, 2022
2022
Closest in time.
L. Campi and M. Fischer, “Correlated equilibria and mean field games: a simple model,” Math. Oper. Res. , 2022
2022
Closest in time.
2022
Closest in time.
H. Gao, W. Lee, Y. Kang, W. Li, Z. Han, S. Osher, and H. V. Poor, “Energy-efficient velocity control for massive numbers of UAVs: A mean field game approach,” IEEE Trans. Veh. Technol. , vol. 71, no. 6, pp. 6266–6278, 2022
2022
Closest in time.
G. Wang, W. Yao, X. Zhang, and Z. Li, “A mean-field game control for large-scale swarm formation flight in dense environments,” Sensors , vol. 22, no. 14, p. 5437, 2022
2022
Closest in time.
W. U. Mondal, M. Agarwal, V. Aggarwal, and S. V. Ukkusuri, “On the approximation of cooperative heterogeneous multi-agent reinforcement learning (MARL) using mean field control (MFC),” J. Mach. Learn. Res. , vol. 23, no. 129, pp. 1–46, 2022
2022
Closest in time.
R. Ourari, K. Cui, A. Elshamanhory, and H. Koeppl, “Nearest-neighbor-based collision avoidance for quadrotors via reinforcement learning,” in Proc. IEEE ICRA , 2022, pp. 293–300
2022
Closest in time.
W. U. Mondal, V. Aggarwal, and S. Ukkusuri, “On the near-optimality of local policies in large cooperative multi-agent reinforcement learning,” Trans. Mach. Learn. Res. , 2022. [Online]. Available: https://openreview.net/forum?id=t5HkgbxZp1
2022
Closest in time.
C. Yu, A. Velu, E. Vinitsky, Y. Wang, A. Bayen, and Y. Wu, “The surprising effectiveness of PPO in cooperative, multi-agent games,” Proc. NeurIPS Datasets and Benchmarks , 2022
2022
Closest in time.
W. Fu, C. Yu, Z. Xu, J. Yang, and Y. Wu, “Revisiting some common practices in cooperative multi-agent reinforcement learning,” in Proc. ICML , 2022, pp. 6863–6877
2022
Closest in time.
K. Cui and H. Koeppl, “Learning graphon mean field games and approximate Nash equilibria,” in Proc. ICLR , 2022, pp. 1–31
2022
Closest in time.
C. Duan, T. Nishikawa, and A. E. Motter, “Prevalence and scalable control of localized networks,” PNAS , vol. 119, no. 32, p. e2122566119, 2022
2022
Closest in time.
S. Perrin, M. Laurière, J. Pérolat, R. Élie, M. Geist, and O. Pietquin, “Generalization in mean field games by learning master policies,” in Proc. AAAI , vol. 36, no. 9, 2022, pp. 9413–9421
2022
Closest in time.