Fetching the paper…
Reading the bibliography…
Robust Reinforcement Learning (RL) focuses on improving performances under model errors or adversarial attacks, which facilitates the real-life deployment of RL agents.
Stochastic games
Lloyd S Shapley · 1953
Earlier work this paper cites.
Gmres: A generalized minimal residual algorithm for solving nonsymmetric linear systems
Youcef Saad and Martin H Schultz · 1986
Earlier work this paper cites.
An introduction to the conjugate gradient method without the agonizing pain, 1994
Jonathan Richard Shewchuk et al · 1994
Earlier work this paper cites.
Dynamic noncooperative game theory
Tamer Başar and Geert Jan Olsder · 1998
Earlier work this paper cites.
Robust reinforcement learning
Jun Morimoto and Kenji Doya · 2005
Earlier work this paper cites.
Manifolds, tensor analysis, and applications
Ralph Abraham, Jerrold E Marsden, and Tudor Ratiu · 2012
Earlier work this paper cites.
Multiple-gradient descent algorithm (mgda) for multiobjective optimization
Jean-Antoine Désidéri · 2012
Earlier work this paper cites.
Estimating the hessian by back-propagating curvature
James Martens, Ilya Sutskever, and Kevin Swersky · 2012
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Learning with opponent-learning awareness
Jakob N Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch · 2017
Earlier work this paper cites.
Robust adversarial reinforcement learning
Lerrel Pinto, James Davidson, Rahul Sukthankar, and Abhinav Gupta · 2017
Cited alongside, same era.
Solving rubik’s cube with a robot hand
Ilge Akkaya, Marcin Andrychowicz, Maciek Chociej, Mateusz Litwin, Bob McGrew, Arthur Petron, Alex Paino, Matthias Plappert, Glenn Powell, Raphael Ribas, et al · 2019
Cited alongside, same era.
Equilibrium strategies for alpha-maxmin expected utility maximization
Bin Li, Peng Luo, and Dewen Xiong · 2019
Cited alongside, same era.
Implicit competitive regularization in gans
Florian Schäfer, Hongkai Zheng, and Anima Anandkumar · 2019
Cited alongside, same era.
Action robust reinforcement learning and applications in continuous control
Chen Tessler, Yonathan Efroni, and Shie Mannor · 2019
Cited alongside, same era.
Deep reinforcement learning with robust and smooth policy
Qianli Shen, Yan Li, Haoming Jiang, Zhaoran Wang, and Tuo Zhao · 2020
Later among the works it cites.
Task-agnostic online reinforcement learning with an infinite mixture of gaussian processes
Mengdi Xu, Wenhao Ding, Jiacheng Zhu, Zuxin Liu, Baiming Chen, and Ding Zhao · 2020
Later among the works it cites.
On the stability and convergence of robust adversarial reinforcement learning: A case study on linear quadratic systems
Kaiqing Zhang, Bin Hu, and Tamer Basar · 2020
Later among the works it cites.
Sim-to-real transfer in deep reinforcement learning for robotics: a survey
Wenshuai Zhao, Jorge Peña Queralta, and Tomi Westerlund · 2020
Later among the works it cites.
Delay-aware model-based reinforcement learning for continuous control
Baiming Chen, Mengdi Xu, Liang Li, and Ding Zhao · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K Zhang, Z Yang, and T Başar · 2019
Cited alongside, same era.
Emergent complexity and zero-shot transfer via unsupervised environment design
Michael Dennis, Natasha Jaques, Eugene Vinitsky, Alexandre Bayen, Stuart Russell, Andrew Critch, and Sergey Levine · 2020
Cited alongside, same era.
Implicit learning dynamics in stackelberg games: Equilibria characterization, convergence analysis, and empirical study
Tanner Fiez, Benjamin Chasnov, and Lillian Ratliff · 2020
Cited alongside, same era.
What is local optimality in nonconvex-nonconcave minimax optimization?
Chi Jin, Praneeth Netrapalli, and Michael Jordan · 2020
Cited alongside, same era.
Robust reinforcement learning via adversarial training with langevin dynamics
Parameswaran Kamalaruban, Yu-Ting Huang, Ya-Ping Hsieh, Paul Rolland, Cheng Shi, and Volkan Cevher · 2020
Cited alongside, same era.
Multimodal safety-critical scenarios generation for decision-making algorithms evaluation
Wenhao Ding, Baiming Chen, Bo Li, Kim Ji Eun, and Ding Zhao · 2021
Later among the works it cites.
Accelerated policy evaluation: Learning adversarial environments with adaptive importance sampling
Mengdi Xu, Peide Huang, Fengpei Li, Jiacheng Zhu, Xuewei Qi, Kentaro Oguchi, Zhiyuan Huang, Henry Lam, and Ding Zhao · 2021
Later among the works it cites.
Robust reinforcement learning: A constrained game-theoretic approach
Jing Yu, Clement Gehring, Florian Schäfer, and Animashree Anandkumar · 2021
Later among the works it cites.
Guodong Zhang, Yuanhao Wang, Laurent Lessard, and Roger Grosse · 2021
Later among the works it cites.
An environment for autonomous driving decision-making
Edouard Leurent · 2022
Closest in time.