Fetching the paper…
Reading the bibliography…
In this paper, we tackle the problem of learning to play 3v3 multi-drone volleyball, a new embodied competitive task that requires both high-level strategic coordination and low-level agile control.
Stochastic games
L. S. Shapley · 1953
Earlier work this paper cites.
Some studies in machine learning using the game of checkers
A. L. Samuel · 1959
Earlier work this paper cites.
The rating of chessplayers: Past and present
A. E. Elo and S. Sloan · 1978
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
M. L. Littman · 1994
Earlier work this paper cites.
Robocup: The robot world cup initiative
H. Kitano, M. Asada, Y. Kuniyoshi, I. Noda, and E. Osawa · 1997
Earlier work this paper cites.
Behavior-based robotics
R. C. Arkin · 1998
Earlier work this paper cites.
Fictitious self-play in extensive-form games
J. Heinrich, M. Lanctot, and D. Silver · 2015
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Earlier work this paper cites.
Rotors—a modular gazebo mav simulator framework
F. Furrer, M. Burri, M. Achtelik, and R. Siegwart · 2016
Earlier work this paper cites.
Hierarchical relative entropy policy search
C. Daniel, G. Neumann, O. Kroemer, and J. Peters · 2016
Earlier work this paper cites.
Deepstack: Expert-level artificial intelligence in heads-up no-limit poker
M. Moravčík, M. Schmid, N. Burch, V. Lisỳ, D. Morrill, N. Bard, T. Davis, K. Waugh, M. Johanson, and M. Bowling · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Cited alongside, same era.
A unified game-theoretic approach to multiagent reinforcement learning
M. Lanctot, V. Zambaldi, A. Gruslys, A. Lazaridou, K. Tuyls, J. Pérolat, D. Silver, and T. Graepel · 2017
Cited alongside, same era.
An algorithmic perspective on imitation learning
T. Osa, J. Pajarinen, G. Neumann, J. A. Bagnell, P. Abbeel, J. Peters, et al · 2018
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, et al · 2019
Cited alongside, same era.
Orbit: A unified simulation framework for interactive robot learning environments
M. Mittal, C. Yu, Q. Yu, J. Liu, N. Rudin, D. Hoeller, J. L. Yuan, R. Singh, Y. Guo, H. Mazhar, et al · 2023
Later among the works it cites.
Champion-level drone racing using deep reinforcement learning
E. Kaufmann, L. Bauersfeld, A. Loquercio, M. Müller, V. Koltun, and D. Scaramuzza · 2023
Later among the works it cites.
Achieving human level competitive robot table tennis
D. B. D’Ambrosio, S. W. Abeyruwan, L. Graesser, A. Iscen, H. B. Amor, A. Bewley, B. Reed, K. Reymann, L. Takayama, Y. Tassa, et al · 2024
Later among the works it cites.
J. Chen, C. Yu, G. Li, W. Tang, X. Yang, B. Xu, H. Yang, and Y. Wang · 2024
Later among the works it cites.
Omnidrones: An efficient and flexible platform for reinforcement learning in drone control
B. Xu, F. Gao, C. Yu, R. Zhang, Y. Wu, and Y. Wang · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Emergent coordination through competition
S. Liu, G. Lever, J. Merel, S. Tunyasuvunakool, N. Heess, and T. Graepel · 2019
Cited alongside, same era.
Isaac gym: High performance gpu-based physics simulation for robot learning
V. Makoviychuk, L. Wawrzyniak, Y. Guo, M. Lu, K. Storey, M. Macklin, D. Hoeller, N. Rudin, A. Allshire, A. Handa, et al · 2021
Cited alongside, same era.
From motor control to team play in simulated humanoid football
S. Liu, G. Lever, Z. Wang, J. Merel, S. A. Eslami, D. Hennes, W. M. Czarnecki, Y. Tassa, S. Omidshafiei, A. Abdolmaleki, et al · 2022
Cited alongside, same era.
Do as i can, not as i say: Grounding language in robotic affordances
M. Ahn, A. Brohan, N. Brown, Y. Chebotar, O. Cortes, B. David, C. Finn, C. Fu, K. Gopalakrishnan, K. Hausman, et al · 2022
Cited alongside, same era.
The surprising effectiveness of ppo in cooperative multi-agent games
C. Yu, A. Velu, E. Vinitsky, J. Gao, Y. Wang, A. Bayen, and Y. Wu · 2022
Cited alongside, same era.
Learning agile soccer skills for a bipedal robot with deep reinforcement learning
T. Haarnoja, B. Moran, G. Lever, S. H. Huang, D. Tirumala, J. Humplik, M. Wulfmeier, S. Tunyasuvunakool, N. Y. Siegel, R. Hafner, et al · 2024
Later among the works it cites.
Mqe: Unleashing the power of interaction with multi-agent quadruped environment
Z. Xiong, B. Chen, S. Huang, W.-W. Tu, Z. He, and Y. Gao · 2024
Later among the works it cites.
A survey on self-play methods in reinforcement learning
R. Zhang, Z. Xu, C. Ma, C. Yu, W.-W. Tu, W. Tang, S. Huang, D. Ye, W. Ding, Y. Yang, et al · 2024
Later among the works it cites.
Mastering table tennis with hierarchy: a reinforcement learning approach with progressive self-play training
H. Ma, J. Fan, H. Xu, and Q. Wang · 2025
Closest in time.
Volleybots: A testbed for multi-drone volleyball game combining motion control and strategic play
Z. Xu, C. Yu, R. Zhang, H. Yuan, X. Yi, S. Ji, C. Wang, W. Tang, and Y. Wang · 2025
Closest in time.