Fetching the paper…
Reading the bibliography…
It is well known that it is difficult to have a reliable and robust framework to link multi-agent deep reinforcement learning algorithms with practical multi-robot applications.
M. J. Matarić, “Reinforcement learning in the multi-robot domain,” in Robot colonies . Springer, 1997, pp. 73–83
1997
Earlier work this paper cites.
O. Michel, “Webots: Symbiosis between virtual and real mobile robots,” in International Conference on Virtual Worlds . Springer, 1998, pp. 254–263
1998
Earlier work this paper cites.
M. A. K. Jaradat, M. Al-Rousan, and L. Quadan, “Reinforcement based mobile robot navigation in dynamic environment,” Robotics and Computer-Integrated Manufacturing , vol. 27, no. 1, pp. 135–149, 2011
2011
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Earlier work this paper cites.
E. Coumans and Y. Bai, “Pybullet, a python module for physics simulation for games, robotics and machine learning,” 2016
2016
Earlier work this paper cites.
Y. Duan, X. Chen, R. Houthooft, J. Schulman, and P. Abbeel, “Benchmarking deep reinforcement learning for continuous control,” 2016
2016
Earlier work this paper cites.
R. Lowe, Y. I. Wu, A. Tamar, J. Harb, O. P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in Advances in neural information processing systems , 2017, pp. 6379–6390
2017
Earlier work this paper cites.
P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, Y. Wu, and P. Zhokhov, “Openai baselines,” https://github.com/openai/baselines
2017
Earlier work this paper cites.
P. Long, T. Fan, X. Liao, W. Liu, H. Zhang, and J. Pan, “Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 6252–6259
2018
Cited alongside, same era.
M. Hüttenrauch, A. Šošić, and G. Neumann, “Local communication protocols for learning complex swarm behaviors with deep reinforcement learning,” in International Conference on Swarm Intelligence . Springer, 2018, pp. 71–83
2018
Cited alongside, same era.
Y. Tassa, Y. Doron, A. Muldal, T. Erez, Y. Li, D. de Las Casas, D. Budden, A. Abdolmaleki, J. Merel, A. Lefrancq, T. Lillicrap, and M. Riedmiller, “Deepmind control suite,” 2018
2018
Cited alongside, same era.
K. S. Krishna, Y. Hao, S. G. Aron, Pradeep, and J. L. Yong, “Hide-and-seek: A data augmentation technique for weakly-supervised localization and beyond,” in Arxiv , 2018
2018
Cited alongside, same era.
O. K. Schulman, “Roboschool,” https://openai.com/blog/roboschool/
2020
Later among the works it cites.
M. A. R. Alberto Ezquerro, “Openai-ros,” http://wiki.ros.org/openai_ros
2020
Later among the works it cites.
M. Kirtas, K. Tsampazis, N. Passalis, and A. Tefas, “Deepbots: A webots-based deep reinforcement learning framework for robotics,” in IFIP International Conference on Artificial Intelligence Applications and Innovations . Springer, 2020, pp. 64–75
2020
Later among the works it cites.
B. Delhaisse, L. Rozo, and D. G. Caldwell, “Pyrobolearn: A python framework for robot learning practitioners,” in Conference on Robot Learning . PMLR, 2020, pp. 1348–1358
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Achiam, “Spinning Up in Deep Reinforcement Learning,” 2018
2018
Cited alongside, same era.
A. Hill, A. Raffin, M. Ernestus, A. Gleave, A. Kanervisto, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Stable baselines,” https://github.com/hill-a/stable-baselines
2018
Cited alongside, same era.
2019
Cited alongside, same era.
Q. Wang, J. Xiong, L. Han, M. Fang, X. Sun, Z. Zheng, P. Sun, and Z. Zhang, “Arena: a toolkit for multi-agent reinforcement learning,” 2019
2019
Cited alongside, same era.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science robotics , vol. 5, no. 47, 2020
2020
Cited alongside, same era.
“https://github.com/unity-technologies/unity-robotics-hub.”
Cited in the paper.
2020
Later among the works it cites.
2020
Later among the works it cites.
A. Juliani, V.-P. Berges, E. Teng, A. Cohen, J. Harper, C. Elion, C. Goy, Y. Gao, H. Henry, M. Mattar, and D. Lange, “Unity: A general platform for intelligent agents,” 2020
2020
Later among the works it cites.
D. Yarats and I. Kostrikov, “Soft actor-critic (sac) implementation in pytorch,” https://github.com/denisyarats/pytorch_sac
2020
Later among the works it cites.
V. Makoviychuk, L. Wawrzyniak, Y. Guo, M. Lu, K. Storey, M. Macklin, D. Hoeller, N. Rudin, A. Allshire, A. Handa, and G. State, “Isaac gym: High performance gpu-based physics simulation for robot learning,” 2021
2021
Later among the works it cites.