Fetching the paper…
Reading the bibliography…
While many multi-robot coordination problems can be solved optimally by exact algorithms, solutions are often not scalable in the number of robots.
In: Proceedings of the 14th annual conference on Computer graphics and interactive techniques, pp. 25–34 (1987)
Reynolds, C.W.: Flocks, herds and schools: A distributed behavioral model · 1987
Earlier work this paper cites.
In: Electrimacs 99 (modelling and simulation of electric machines converters an& systems), pp. I–71 (1999)
Niiranen, J.: Fast and accurate symmetric euler algorithm for electromechanical simulations · 1999
Earlier work this paper cites.
Autonomous Robots 11
Ijspeert, A.J., Martinoli, A., Billard, A., Gambardella, L.M.: Collaboration through the exploitation of local interactions in autonomous collective robotics: The stick pulling experiment · 2001
Earlier work this paper cites.
Mathematics of operations research 27
Bernstein, D.S., Givan, R., Immerman, N., Zilberstein, S.: The complexity of decentralized control of markov decision processes · 2002
Earlier work this paper cites.
In: 2004 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)(IEEE Cat. No. 04CH37566), pp. 2149–2154. IEEE (2004)
Koenig, N., Howard, A.: Design and use paradigms for gazebo, an open-source multi-robot simulator · 2004
Earlier work this paper cites.
International Journal of Advanced Robotic Systems p. 5 (2004)
Michel, O.: Cyberbotics ltd. webots™: professional mobile robot simulation · 2004
Earlier work this paper cites.
Transportation science 39
Bräysy, O., Gendreau, M.: Vehicle routing problem with time windows, part ii: Metaheuristics · 2005
Earlier work this paper cites.
IEEE Transactions on Robotics pp. 1018–1031 (2010)
Zheng, X., Koenig, S., Kempe, D., Jain, S.: Multirobot forest coverage for weighted and unweighted terrain · 2010
Earlier work this paper cites.
Swarm Intelligence pp. 271–295 (2012)
Pinciroli, C., Trianni, V., O’Grady, R., Pini, G., Brutschy, A., Brambilla, M., Mathews, N., Ferrante, E., Di Caro, G., Ducatelle, F., Birattari, M., Gambardella, L.M., Dorigo, M.: ARGoS: a modular, parallel, multi-engine simulator for multi-robot systems · 2012
Earlier work this paper cites.
In: 2012 IEEE/RSJ international conference on intelligent robots and systems, pp. 5026–5033. IEEE (2012)
Todorov, E., Erez, T., Tassa, Y.: Mujoco: A physics engine for model-based control · 2012
Earlier work this paper cites.
arXiv preprint arXiv:1606.01540 (2016)
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., Zaremba, W.: Openai gym · 2016
Earlier work this paper cites.
Advances in neural information processing systems (2017)
Lowe, R., Wu, Y.I., Tamar, A., Harb, J., Pieter Abbeel, O., Mordatch, I.: Multi-agent actor-critic for mixed cooperative-competitive environments · 2017
Earlier work this paper cites.
In: 2017 IEEE International Symposium on Safety, Security and Rescue Robotics (SSRR), pp. 19–24 (2017)
Noori, F.M., Portugal, D., Rocha, R.P., Couceiro, M.S.: On 3d simulators for multi-robot systems in ros: Morse or gazebo? · 2017
Earlier work this paper cites.
arXiv preprint arXiv:1707.06347 (2017)
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., Klimov, O.: Proximal policy optimization algorithms · 2017
Cited alongside, same era.
URL http://github.com/google/jax
Bradbury, J., Frostig, R., Hawkins, P., Johnson, M.J., Leary, C., Maclaurin, D., Necula, G., Paszke, A., VanderPlas, J., Wanderman-Milne, S., Zhang, Q.: JAX: composable transformations of Python+NumPy programs (2018) · 2018
Cited alongside, same era.
In: International Conference on Machine Learning, pp. 3053–3062. PMLR (2018)
Liang, E., Liaw, R., Nishihara, R., Moritz, P., Fox, R., Goldberg, K., Gonzalez, J., Jordan, M., Stoica, I.: Rllib: Abstractions for distributed reinforcement learning · 2018
Cited alongside, same era.
CoRR (2018)
Resnick, C., Eldridge, W., Ha, D., Britz, D., Foerster, J., Togelius, J., Cho, K., Bruna, J.: Pommerman: A multi-agent playground · 2018
Cited alongside, same era.
In: Proceedings of the AAAI conference on artificial intelligence (2018)
Zheng, L., Yang, J., Cai, H., Zhou, M., Zhang, W., Wang, J., Yu, Y.: Magent: A many-agent reinforcement learning platform for artificial collective intelligence · 2018
IEEE Robotics and Automation Letters pp. 6932–6939 (2020)
Wang, B., Liu, Z., Li, Q., Prorok, A.: Mobile robot path planning in dynamic environments through globally guided reinforcement learning · 2020
Later among the works it cites.
arXiv preprint arXiv:2011.09533 (2020)
de Witt, C.S., Gupta, T., Makoviichuk, D., Makoviychuk, V., Torr, P.H., Sun, M., Whiteson, S.: Is independent learning all you need in the starcraft multi-agent challenge? · 2020
Later among the works it cites.
arXiv preprint arXiv:2111.01777 (2021)
Blumenkamp, J., Morad, S., Gielis, J., Li, Q., Prorok, A.: A framework for real-world multi-robot systems running decentralized gnn-based policies · 2021
Later among the works it cites.
URL http://github.com/google/brax
Freeman, C.D., Frey, E., Raichuk, A., Girgin, S., Mordatch, I., Bachem, O.: Brax - a differentiable physics engine for large scale rigid body simulation (2021) · 2021
Later among the works it cites.
In: Proceedings of the 36th Annual ACM Symposium on Applied Computing, pp. 777–784 (2021)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
In: International Conference on Learning Representations (2019)
Baker, B., Kanitscheider, I., Markov, T., Wu, Y., Powell, G., McGrew, B., Mordatch, I.: Emergent tool use from multi-agent autocurricula · 2019
Cited alongside, same era.
Advances in neural information processing systems (2019)
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., et al.: Pytorch: An imperative style, high-performance deep learning library · 2019
Cited alongside, same era.
CoRR (2019)
Samvelyan, M., Rashid, T., de Witt, C.S., Farquhar, G., Nardelli, N., Rudner, T.G.J., Hung, C.M., Torr, P.H.S., Foerster, J., Whiteson, S.: The StarCraft Multi-Agent Challenge · 2019
Cited alongside, same era.
arXiv preprint arXiv:1903.00784 (2019)
Suarez, J., Du, Y., Isola, P., Mordatch, I.: Neural mmo: A massively multiagent game environment for training and evaluating intelligent agents · 2019
Cited alongside, same era.
In: Proceedings of the AAAI Conference on Artificial Intelligence, pp. 4501–4510 (2020)
Kurach, K., Raichuk, A., Stańczyk, P., Zając, M., Bachem, O., Espeholt, L., Riquelme, C., Vincent, D., Michalski, M., Bousquet, O., et al.: Google research football: A novel reinforcement learning environment · 2020
Cited alongside, same era.
In: 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), pp. 11,785–11,792. IEEE (2020)
Li, Q., Gama, F., Ribeiro, A., Prorok, A.: Graph neural networks for decentralized multi-robot path planning · 2020
Cited alongside, same era.
IEEE Transactions on Automation Science and Engineering pp. 2025–2037 (2020)
Prorok, A.: Robust assignment using redundant robots on transport networks with uncertain travel time · 2020
Cited alongside, same era.
Jiang, S., Amato, C.: Multi-agent reinforcement learning with directed exploration and selective memory reuse · 2021
Later among the works it cites.
In: Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2) (2021)
Makoviychuk, V., Wawrzyniak, L., Guo, Y., Lu, M., Storey, K., Macklin, M., Hoeller, D., Rudin, N., Allshire, A., Handa, A., et al.: Isaac gym: High performance gpu based physics simulation for robot learning · 2021
Later among the works it cites.
In: 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), pp. 7512–7519. IEEE (2021)
Panerati, J., Zheng, H., Zhou, S., Xu, J., Prorok, A., Schoellig, A.P.: Learning to fly—a gym environment with pybullet physics for reinforcement learning of multi-agent quadcopter control · 2021
Later among the works it cites.
Advances in Neural Information Processing Systems pp. 12,208–12,221 (2021)
Peng, B., Rashid, T., Schroeder de Witt, C., Kamienny, P.A., Torr, P., Böhmer, W., Whiteson, S.: Facmac: Factored multi-agent centralised policy gradients · 2021
Later among the works it cites.
arXiv preprint arXiv:2103.01955 (2021)
Yu, C., Velu, A., Vinitsky, E., Wang, Y., Bayen, A., Wu, Y.: The surprising effectiveness of ppo in cooperative, multi-agent games · 2021
Later among the works it cites.
URL http://github.com/RobertTLange/gymnax
Lange, R.T.: gymnax: A JAX-based reinforcement learning environment library (2022) · 2022
Closest in time.
arXiv preprint arXiv:2203.06464 (2022)
Shen, J., Xiao, E., Liu, Y., Feng, C.: A deep reinforcement learning environment for particle robot navigation and object manipulation · 2022
Closest in time.
arXiv preprint arXiv:2206.10558 (2022)
Weng, J., Lin, M., Huang, S., Liu, B., Makoviichuk, D., Makoviychuk, V., Liu, Z., Song, Y., Luo, T., Jiang, Y., Xu, Z., Yan, S.: Envpool: A highly parallel reinforcement learning environment execution engine · 2022
Closest in time.