Fetching the paper…
Reading the bibliography…
Reinforcement learning has achieved remarkable performance in a wide range of tasks these days.
K. J. Astrom, “Optimal control of markov processes with incomplete state information,” Journal of mathematical analysis and applications , vol. 10, no. 1, pp. 174–205, 1965
1965
Earlier work this paper cites.
G. N. Iyengar, “Robust dynamic programming,” Mathematics of Operations Research , vol. 30, no. 2, pp. 257–280, 2005
2005
Earlier work this paper cites.
Y. Bengio and Y. LeCun, “Scaling learning algorithms towards AI,” in Large Scale Kernel Machines . MIT Press, 2007
2007
Earlier work this paper cites.
B. D. Ziebart, “Modeling purposeful adaptive behavior with the principle of maximum causal entropy,” 2010
2010
Earlier work this paper cites.
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al. , “Mastering the game of go with deep neural networks and tree search,” nature , vol. 529, no. 7587, pp. 484–489, 2016
2016
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning,” in International Conference on Machine Learning , 2017, pp. 2817–2826
2017
Cited alongside, same era.
A. Mandlekar, Y. Zhu, A. Garg, L. Fei-Fei, and S. Savarese, “Adversarially robust policy learning: Active construction of physically-plausible perturbations,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2017, pp. 3932–3939
2017
Cited alongside, same era.
A. Roy, H. Xu, and S. Pokutta, “Reinforcement learning under model mismatch,” in Advances in neural information processing systems , 2017, pp. 3043–3052
2017
Cited alongside, same era.
S. Gu, E. Holly, T. Lillicrap, and S. Levine, “Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates,” in 2017 IEEE international conference on robotics and automation (ICRA) . IEEE, 2017, pp. 3389–3396
2017
Cited alongside, same era.
A. Hill, A. Raffin, M. Ernestus, A. Gleave, A. Kanervisto, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Stable baselines,” https://github.com/hill-a/stable-baselines , 2018
2018
Later among the works it cites.
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, et al. , “Grandmaster level in starcraft ii using multi-agent reinforcement learning,” Nature , vol. 575, no. 7782, pp. 350–354, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
D. J. Mankowitz, N. Levine, R. Jeong, A. Abdolmaleki, J. T. Springenberg, Y. Shi, J. Kay, T. Hester, T. Mann, and M. Riedmiller, “Robust reinforcement learning for continuous control with model misspecification,” in International Conference on Learning Representations , 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Finn and S. Levine, “Deep visual foresight for planning robot motion,” in 2017 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2017, pp. 2786–2793
2017
Cited alongside, same era.
T. Haarnoja, H. Tang, P. Abbeel, and S. Levine, “Reinforcement learning with deep energy-based policies,” in International Conference on Machine Learning , 2017, pp. 1352–1361
2017
Cited alongside, same era.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 1–8
2018
Cited alongside, same era.
2018
Cited alongside, same era.
A. Havens, Z. Jiang, and S. Sarkar, “Online robust policy learning in the presence of unknown adversaries,” in Advances in Neural Information Processing Systems , 2018, pp. 9916–9926
2018
Cited alongside, same era.
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International Conference on Machine Learning , 2018, pp. 1861–1870
2018
Cited alongside, same era.
2019
Later among the works it cites.
X. Pan, D. Seita, Y. Gao, and J. Canny, “Risk averse robust adversarial reinforcement learning,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 8522–8528
2019
Later among the works it cites.
2019
Later among the works it cites.
C. Tessler, Y. Efroni, and S. Mannor, “Action robust reinforcement learning and applications in continuous control,” in International Conference on Machine Learning , 2019, pp. 6215–6224
2019
Later among the works it cites.
C. Wang, J. Wang, Y. Shen, and X. Zhang, “Autonomous navigation of uavs in large-scale complex environments: A deep reinforcement learning approach,” IEEE Transactions on Vehicular Technology , vol. 68, no. 3, pp. 2124--2136, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
A. Pattanaik, Z. Tang, S. Liu, G. Bommannan, and G. Chowdhary, “Robust deep reinforcement learning with adversarial attacks,” in Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems , 2018, pp. 2040–2042
2042
Closest in time.