Fetching the paper…
Reading the bibliography…
We study the robustness of reinforcement learning (RL) with adversarially perturbed state observations, which aligns with the setting of many adversarial attacks to deep reinforcement learning (DRL) and is also important for rolling out real-world RL agent under unpredictable sensing noise.
Optimal control of markov processes with incomplete state information
Karl J Astrom · 1965
Earlier work this paper cites.
Artificial life and real robots
Rodney A Brooks · 1992
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Michael L Littman · 1994
Earlier work this paper cites.
Noise and the reality gap: The use of simulation in evolutionary robotics
Nick Jakobi, Phil Husbands, and Inman Harvey · 1995
Earlier work this paper cites.
Robustness in Markov decision problems with uncertain transition matrices
Arnab Nilim and Laurent El Ghaoui · 2004
Earlier work this paper cites.
Robust dynamic programming
Garud N Iyengar · 2005
Earlier work this paper cites.
Solving deep memory pomdps with recurrent policy gradients
Daan Wierstra, Alexander Foerster, Jan Peters, and Juergen Schmidhuber · 2007
Earlier work this paper cites.
Sarsop: Efficient point-based pomdp planning by approximating optimally reachable belief spaces
Hanna Kurniawati, David Hsu, and Wee Sun Lee · 2008
Earlier work this paper cites.
Value-based policy teaching with active indirect elicitation
Haoqi Zhang and David C Parkes · 2008
Earlier work this paper cites.
Policy teaching through reward function learning
Haoqi Zhang, David C Parkes, and Yiling Chen · 2009
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2013
Earlier work this paper cites.
Integrated perception and planning in the continuous space: A pomdp approach
Haoyu Bai, David Hsu, and Wee Sun Lee · 2014
Earlier work this paper cites.
Deep learning for real-time atari game play using offline monte-carlo tree search planning
Xiaoxiao Guo, Satinder Singh, Honglak Lee, Richard L Lewis, and Xiaoshi Wang · 2014
Earlier work this paper cites.
Markov decision processes: discrete stochastic dynamic programming
Martin L Puterman · 2014
Earlier work this paper cites.
Deep recurrent q-learning for partially observable mdps
Matthew Hausknecht and Peter Stone · 2015
Earlier work this paper cites.
Trust region policy optimization
Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Earlier work this paper cites.
Distributional smoothing with virtual adversarial training
Takeru Miyato, Shin-ichi Maeda, Masanori Koyama, Ken Nakae, and Shin Ishii · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Robust partially observable Markov decision process
Takayuki Osogami · 2015
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Cited alongside, same era.
Adversarial attacks on neural network policies
Sandy Huang, Nicolas Papernot, Ian Goodfellow, Yan Duan, and Pieter Abbeel · 2017
Cited alongside, same era.
Delving into adversarial attacks on deep policies
Jernej Kos and Dawn Song · 2017
Cited alongside, same era.
Tactics of adversarial attack on deep reinforcement learning agents
Yen-Chen Lin, Zhang-Wei Hong, Yuan-Hong Liao, Meng-Li Shih, Ming-Yu Liu, and Min Sun · 2017
Online robustness training for deep reinforcement learning
Marc Fischer, Matthew Mirman, and Martin Vechev · 2019
Later among the works it cites.
Adversarial policies: Attacking deep reinforcement learning
Adam Gleave, Michael Dennis, Neel Kant, Cody Wild, Sergey Levine, and Stuart Russell · 2019
Later among the works it cites.
Deceptive reinforcement learning under adversarial manipulations on cost signals
Yunhan Huang and Quanyan Zhu · 2019
Later among the works it cites.
Snooping attacks on deep reinforcement learning
Matthew Inkawhich, Yiran Chen, and Hai Li · 2019
Later among the works it cites.
Robust multi-agent reinforcement learning via minimax deep deterministic policy gradient
Shihui Li, Yi Wu, Xinyue Cui, Honghua Dong, Fei Fang, and Stuart Russell · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Adversarially robust policy learning: Active construction of physically-plausible perturbations
Ajay Mandlekar, Yuke Zhu, Animesh Garg, Li Fei-Fei, and Silvio Savarese · 2017
Cited alongside, same era.
Robust adversarial reinforcement learning
Lerrel Pinto, James Davidson, Rahul Sukthankar, and Abhinav Gupta · 2017
Cited alongside, same era.
Deep reinforcement learning framework for autonomous driving
Ahmad EL Sallab, Mohammed Abdou, Etienne Perot, and Senthil Yogamani · 2017
Cited alongside, same era.
Online algorithms for pomdps with continuous state, action, and observation spaces
Zachary Sunberg and Mykel Kochenderfer · 2017
Cited alongside, same era.
Trust region policy optimization for pomdps
Kamyar Azizzadenesheli, Manish Kumar Bera, and Animashree Anandkumar · 2018
Cited alongside, same era.
Addressing function approximation error in actor-critic methods
Scott Fujimoto, Herke Van Hoof, and David Meger · 2018
Cited alongside, same era.
On the effectiveness of interval bound propagation for training verifiably robust models
Sven Gowal, Krishnamurthy Dvijotham, Robert Stanforth, Rudy Bunel, Chongli Qin, Jonathan Uesato, Timothy Mann, and Pushmeet Kohli · 2018
Cited alongside, same era.
Later among the works it cites.
Policy poisoning in batch reinforcement learning and control
Yuzhe Ma, Xuezhou Zhang, Wen Sun, and Jerry Zhu · 2019
Later among the works it cites.
Robust reinforcement learning for continuous control with model misspecification
Daniel J Mankowitz, Nir Levine, Rae Jeong, Abbas Abdolmaleki, Jost Tobias Springenberg, Timothy Mann, Todd Hester, and Martin Riedmiller · 2019
Later among the works it cites.
Assessing transferability from simulation to reality for reinforcement learning
Fabio Muratore, Michael Gienger, and Jan Peters · 2019
Later among the works it cites.
A convex relaxation barrier to tight robustness verification of neural networks
Hadi Salman, Greg Yang, Huan Zhang, Cho-Jui Hsieh, and Pengchuan Zhang · 2019
Later among the works it cites.
Action robust reinforcement learning and applications in continuous control
Chen Tessler, Yonathan Efroni, and Shie Mannor · 2019
Later among the works it cites.
Introducing voyage deepdrive unlocking the potential of deep reinforcement learning
Voyage · 2019
Later among the works it cites.
Characterizing attacks on deep reinforcement learning
Chaowei Xiao, Xinlei Pan, Warren He, Jian Peng, Mingjie Sun, Jinfeng Yi, Bo Li, and Dawn Song · 2019
Later among the works it cites.
Theoretically principled trade-off between robustness and accuracy
Hongyang Zhang, Yaodong Yu, Jiantao Jiao, Eric P Xing, Laurent El Ghaoui, and Michael I Jordan · 2019
Later among the works it cites.
Implementation matters in deep policy gradients: A case study on PPO and TRPO
Logan Engstrom, Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras, Firdaus Janoos, Larry Rudolph, and Aleksander Madry · 2020
Later among the works it cites.
Defending adversarial attacks without adversarial attacks in deep reinforcement learning
Xinghua Qu, Yew-Soon Ong, Abhishek Gupta, and Zhu Sun · 2020
Later among the works it cites.
Amin Rakhsha, Goran Radanovic, Rati Devidze, Xiaojin Zhu, and Adish Singla · 2020
Later among the works it cites.
Robustifying reinforcement learning agents via action space adversarial training
Kai Liang Tan, Yasaman Esfandiari, Xian Yeow Lee, Soumik Sarkar, et al · 2020
Later among the works it cites.
Automatic perturbation analysis on general computational graphs
Kaidi Xu, Zhouxing Shi, Huan Zhang, Minlie Huang, Kai-Wei Chang, Bhavya Kailkhura, Xue Lin, and Cho-Jui Hsieh · 2020
Later among the works it cites.