Fetching the paper…
Reading the bibliography…
Learning strategic robot behavior -- like that required in pursuit-evasion interactions -- under real-world constraints is extremely challenging.
Stochastic games
L. S. Shapley · 1953
Earlier work this paper cites.
Differential games i: Introduction
R. Isaacs · 1954
Earlier work this paper cites.
On curves of minimal length with a constraint on average curvature, and with prescribed initial and terminal positions and tangents
L. E. Dubins · 1957
Earlier work this paper cites.
A new approach to linear filtering and prediction problems
R. E. Kalman · 1960
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
M. L. Littman · 1994
Earlier work this paper cites.
Dynamic noncooperative game theory
T. Başar and G. J. Olsder · 1998
Earlier work this paper cites.
Differential games: a mathematical theory with applications to warfare and pursuit, control and optimization
R. Isaacs · 1999
Earlier work this paper cites.
A time-dependent hamilton-jacobi formulation of reachable sets for continuous dynamic games
I. M. Mitchell, A. M. Bayen, and C. J. Tomlin · 2005
Earlier work this paper cites.
The flexible, extensible and efficient toolbox of level set methods
I. M. Mitchell · 2008
Earlier work this paper cites.
Multi-agent reinforcement learning: An overview
L. Buşoniu, R. Babuška, and B. De Schutter · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
S. Ross, G. Gordon, and D. Bagnell · 2011
Earlier work this paper cites.
Locomotion dynamics of hunting in wild cheetahs
A. M. Wilson, J. Lowe, K. Roskilly, P. E. Hudson, K. Golabek, and J. McNutt · 2013
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Earlier work this paper cites.
Agile autonomous driving using end-to-end deep imitation learning
Y. Pan, C.-A. Cheng, K. Saigol, K. Lee, X. Yan, E. Theodorou, and B. Boots · 2017
Earlier work this paper cites.
Robust adversarial reinforcement learning
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta · 2017
Earlier work this paper cites.
Multi-agent actor-critic for mixed cooperative-competitive environments
R. Lowe, Y. I. Wu, A. Tamar, J. Harb, O. Pieter Abbeel, and I. Mordatch · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Cited alongside, same era.
Counterfactual multi-agent policy gradients
J. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson · 2018
Cited alongside, same era.
Emergent tool use from multi-agent interaction
B. Baker, I. Kanitscheider, T. Markov, Y. Wu, G. Powell, B. McGrew, and I. Mordatch · 2019
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, et al · 2019
Cited alongside, same era.
Multi-agent manipulation via locomotion using hierarchical sim2real
O. Nachum, M. Ahn, H. Ponte, S. Gu, and V. Kumar · 2019
Cited alongside, same era.
Iterative residual policy: for goal-conditioned dynamic manipulation of deformable objects
C. Chi, B. Burchfiel, E. Cousineau, S. Feng, and S. Song · 2022
Later among the works it cites.
Multi-agent deep reinforcement learning: a survey
S. Gronauer and K. Diepold · 2022
Later among the works it cites.
Isaacs: Iterative soft adversarial actor-critic for safety
K.-C. Hsu, D. P. Nguyen, and J. F. Fisac · 2022
Later among the works it cites.
Influencing long-term behavior in multiagent reinforcement learning
D.-K. Kim, M. Riemer, M. Liu, J. Foerster, M. Everett, C. Sun, G. Tesauro, and J. P. How · 2022
Later among the works it cites.
Human-level play in the game of diplomacy by combining language models with strategic reasoning
M. F. A. R. D. T. (FAIR)†, A. Bakhtin, N. Brown, E. Dinan, G. Farina, C. Flaherty, D. Fried, A. Goff, J. Gray, H. Hu, et al · 2022
Later among the works it cites.
Rili: Robustly influencing latent intent
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter · 2020
Cited alongside, same era.
Learning by cheating
D. Chen, B. Zhou, V. Koltun, and P. Krähenbühl · 2020
Cited alongside, same era.
Learning agile locomotion via adversarial training
Y. Tang, J. Tan, and T. Harada · 2020
Cited alongside, same era.
Learning to interactively learn and assist
M. Woodward, C. Finn, and K. Hausman · 2020
Cited alongside, same era.
Multimodal sensor fusion with differentiable filters
M. A. Lee, B. Yi, R. Martín-Martín, S. Savarese, and J. Bohg · 2020
Cited alongside, same era.
Rma: Rapid motor adaptation for legged robots
A. Kumar, Z. Fu, D. Pathak, and J. Malik · 2021
Cited alongside, same era.
Learning high-speed flight in the wild
A. Loquercio, E. Kaufmann, R. Ranftl, M. Müller, V. Koltun, and D. Scaramuzza · 2021
Cited alongside, same era.
S. Parekh, S. Habibian, and D. P. Losey · 2022
Later among the works it cites.
Influencing towards stable multi-agent interactions
W. Z. Wang, A. Shih, A. Xie, and D. Sadigh · 2022
Later among the works it cites.
Creating a dynamic quadrupedal robotic goalkeeper with reinforcement learning
X. Huang, Z. Li, Y. Xiang, Y. Ni, Y. Chi, Y. Li, L. Yang, X. B. Peng, and K. Sreenath · 2022
Later among the works it cites.
Distributed data-driven predictive control for multi-agent collaborative legged locomotion
R. T. Fawcett, L. Amanzadeh, J. Kim, A. D. Ames, and K. A. Hamed · 2022
Later among the works it cites.
Learning to walk in minutes using massively parallel deep reinforcement learning
N. Rudin, D. Hoeller, P. Reist, and M. Hutter · 2022
Later among the works it cites.
Fast traversability estimation for wild visual navigation
J. Frey, M. Mattamala, N. Chebrolu, C. Cadena, M. Fallon, and M. Hutter · 2023
Closest in time.
Learning Visual Locomotion with Cross-Modal Supervision
A. Loquercio, A. Kumar, and J. Malik · 2023
Closest in time.
Learning representations that enable generalization in assistive tasks
J. Z.-Y. He, Z. Erickson, D. S. Brown, A. Raghunathan, and A. Dragan · 2023
Closest in time.
Safety-critical coordination for cooperative legged locomotion via control barrier functions
J. Kim, J. Lee, and A. D. Ames · 2023
Closest in time.
https://www.stereolabs.com/zed-2/
Zed 2 camera · 2023
Closest in time.