Fetching the paper…
Reading the bibliography…
Deep Reinforcement Learning (DRL) has numerous applications in the real world thanks to its outstanding ability in quickly adapting to the surrounding environments.
M. Wiering, “Multi-agent reinforcement learning for traffic light control,” in Machine Learning: Proceedings of the Seventeenth International Conference (ICML’2000) , 2000, pp. 1151–1158
2000
Earlier work this paper cites.
2001
Earlier work this paper cites.
Y. S. Abu-Mostafa, M. Magdon-Ismail, and H.-T. Lin, Learning from data . AMLBook New York, NY, USA, 2012, vol. 4
2012
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, p. 529, Jun. 2015
2015
Earlier work this paper cites.
Abadi et al. , “TensorFlow: Large-scale machine learning on heterogeneous systems,” 2015, software available from tensorflow.org. [Online]. Available: https://www.tensorflow.org/
2015
Earlier work this paper cites.
Y. Deng, F. Bao, Y. Kong, Z. Ren, and Q. Dai, “Deep direct reinforcement learning for financial signal representation and trading,” IEEE transactions on neural networks and learning systems , vol. 28, no. 3, pp. 653–664, Feb. 2016
2016
Earlier work this paper cites.
V. François-Lavet, D. Taralla, D. Ernst, and R. Fonteneau, “Deep reinforcement learning solutions for energy microgrids management,” in European Workshop on Reinforcement Learning (EWRL) , Dec. 2016
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of Go with deep neural networks and tree search,” Nature , vol. 529, no. 7587, p. 484, Dec. 2016
2016
Earlier work this paper cites.
N. Papernot, P. McDaniel, S. Jha, M. Fredrikson, Z. B. Celik, and A. Swami, “The limitations of deep learning in adversarial settings,” in Security and Privacy (EuroS&P), 2016 IEEE European Symposium on . IEEE, Mar. 2016, pp. 372–387
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
I. Osband, C. Blundell, A. Pritzel, and B. Van Roy, “Deep exploration via bootstrapped DQN,” in Advances in neural information processing systems , 2016, pp. 4026–4034
2016
Earlier work this paper cites.
N. Papernot, P. McDaniel, X. Wu, S. Jha, and A. Swami, “Distillation as a defense to adversarial perturbations against deep neural networks,” in 2016 IEEE Symposium on Security and Privacy (SP) . IEEE, May 2016, pp. 582–597
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. A. Rusu, S. G. Colmenarejo, C. Gulcehre, G. Desjardins, J. Kirkpatrick, R. Pascanu, V. Mnih, K. Kavukcuoglu, and R. Hadsell, “Policy distillation,” ICLR , May 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, Nov. 2017
2017
Earlier work this paper cites.
S. Gu, E. Holly, T. Lillicrap, and S. Levine, “Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates,” in IEEE international conference on robotics and automation (ICRA) . IEEE, May 2017, pp. 3389–3396
2017
Earlier work this paper cites.
X. B. Peng, G. Berseth, K. Yin, and M. Van De Panne, “Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning,” ACM Transactions on Graphics (TOG) , vol. 36, no. 4, p. 41, Dec. 2017
2017
Earlier work this paper cites.
A. Raghu, M. Komorowski, L. A. Celi, P. Szolovits, and M. Ghassemi, “Continuous state-space models for optimal sepsis treatment-a deep reinforcement learning approach,” Proceedings of Machine Learning for Healthcare , Apr 2017
2017
Earlier work this paper cites.
C. N. Madu, C.-h. Kuei, and P. Lee, “Urban sustainability management: A deep learning perspective,” Sustainable Cities and Society , vol. 30, pp. 1–17, 2017
2017
Earlier work this paper cites.
V. Behzadan and A. Munir, “Vulnerability of deep reinforcement learning to policy induction attacks,” in International Conference on Machine Learning and Data Mining in Pattern Recognition . Springer, Jul. 2017, pp. 262–275
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
P.-Y. Chen, H. Zhang, Y. Sharma, J. Yi, and C.-J. Hsieh, “ZOO: Zeroth order optimization based black-box attacks to deep neural networks without training substitute models,” in Proceedings of the 10th ACM Workshop on Artificial Intelligence and Security , 2017, pp. 15–26
2017
Earlier work this paper cites.
S. Huang, N. Papernot, I. Goodfellow, Y. Duan, and P. Abbeel, “Adversarial attacks on neural network policies,” ICLR Workshop , 3 2017
2017
Earlier work this paper cites.
Y.-C. Lin, Z.-W. Hong, Y.-H. Liao, M.-L. Shih, M.-Y. Liu, and M. Sun, “Tactics of adversarial attack on deep reinforcement learning agents,” International Joint Conferences on Artificial Intelligence , 8 2017
2017
Earlier work this paper cites.
N. Carlini and D. Wagner, “Towards evaluating the robustness of neural networks,” in Security and Privacy (SP), 2017 IEEE Symposium on . IEEE, May 2017, pp. 39–57
2017
Earlier work this paper cites.
J. Kos and D. Song, “Delving into adversarial attacks on deep policies,” ICLR Workshop , Apr. 2017
2017
Earlier work this paper cites.
X. Pan, Y. You, Z. Wang, and C. Lu, “Virtual to real reinforcement learning for autonomous driving,” Proceedings of the British Machine Vision Conference (BMVC) , Sep. 2017
2017
Earlier work this paper cites.
Z. Wang, V. Bapst, N. Heess, V. Mnih, R. Munos, K. Kavukcuoglu, and N. de Freitas, “Sample efficient actor-critic with experience replay,” ICLR , 2017
2017
Earlier work this paper cites.
Y. Wu, E. Mansimov, R. B. Grosse, S. Liao, and J. Ba, “Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation,” in Advances in neural information processing systems , 2017, pp. 5279–5288
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning-Volume 70 . JMLR. org, Aug. 2017, pp. 2817–2826
2017
Earlier work this paper cites.
M. Bravo and P. Mertikopoulos, “On the robustness of learning in games with stochastically perturbed payoff observations,” Games and Economic Behavior , vol. 103, pp. 41–66, May 2017
2017
Earlier work this paper cites.
A. Mandlekar, Y. Zhu, A. Garg, L. Fei-Fei, and S. Savarese, “Adversarially robust policy learning: Active construction of physically-plausible perturbations,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, Sep. 2017, pp. 3932–3939
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, Y. Wu, and P. Zhokhov, “OpenAI baselines,” https://github.com/openai/baselines , 2017
2017
Earlier work this paper cites.
I. Caspi, G. Leibovich, G. Novik, and S. Endrawis, “Reinforcement learning coach,” Dec 2017. [Online]. Available: https://doi.org/10.5281/zenodo.1134899
2017
Earlier work this paper cites.
S.-M. Moosavi-Dezfooli, A. Fawzi, O. Fawzi, and P. Frossard, “Universal adversarial perturbations,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , July 2017
2017
Cited alongside, same era.
A. R. Fayjie, S. Hossain, D. Oualid, and D.-J. Lee, “Driverless car: Autonomous driving using deep reinforcement learning in urban environment,” in 15th International Conference on Ubiquitous Robots (UR) . IEEE, Jun. 2018, pp. 896–901
2018
Cited alongside, same era.
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel et al. , “A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,” Science , vol. 362, no. 6419, pp. 1140–1144, 2018
2018
Cited alongside, same era.
N. Akhtar and A. Mian, “Threat of adversarial attacks on deep learning in computer vision: A survey,” IEEE Access , vol. 6, pp. 14 410–14 430, 2018
2018
Cited alongside, same era.
X. Pan, D. Seita, Y. Gao, and J. Canny, “Risk averse robust adversarial reinforcement learning,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 8522–8528
2019
Later among the works it cites.
V. Gallego, R. Naveiro, and D. R. Insua, “Reinforcement learning under threats,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 33, Jul. 2019, pp. 9939–9940
2019
Later among the works it cites.
W. M. Czarnecki, R. Pascanu, S. Osindero, S. Jayakumar, G. Swirszcz, and M. Jaderberg, “Distilling policy distillation,” in Proceedings of Machine Learning Research , Apr 2019, pp. 1331–1340
2019
Later among the works it cites.
V. Behzadan and A. Munir, “Adversarial reinforcement learning framework for benchmarking collision avoidance mechanisms in autonomous vehicles,” IEEE Intelligent Transportation Systems Magazine , Apr 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Y. Li, “Deep reinforcement learning,” arXiv preprint arXiv:1810.06339 , Oct. 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
G. Clark, M. Doran, and W. Glisson, “A malicious attack on the machine learning policy of a robotic system,” in 2018 17th IEEE International Conference On Trust, Security And Privacy In Computing And Communications/12th IEEE International Conference On Big Data Science And Engineering (TrustCom/BigDataSE) . IEEE, Aug. 2018, pp. 516–521
2018
Cited alongside, same era.
Y. Vorobeychik and M. Kantarcioglu, “Adversarial machine learning,” Synthesis Lectures on Artificial Intelligence and Machine Learning , vol. 12, no. 3, pp. 1–169, Jun. 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
S. Baluja and I. Fischer, “Learning to attack: Adversarial transformation networks,” in Association for the Advancement of Artificial Intelligence , Feb. 2018, pp. 2687–2695
2018
Cited alongside, same era.
2018
Cited alongside, same era.
V. Behzadan and W. Hsu, “RL-based method for benchmarking the adversarial resilience and robustness of deep reinforcement learning policies,” in International Conference on Computer Safety, Reliability, and Security . Springer, 2019, pp. 314–325
2019
Later among the works it cites.
2019
Later among the works it cites.
P. Gawłowicz and A. Zubow, “ns-3 meets OpenAI Gym: The Playground for Machine Learning in Networking Research,” in ACM International Conference on Modeling, Analysis and Simulation of Wireless and Mobile Systems (MSWiM) , 11 2019. [Online]. Available: http://www.tkn.tu-berlin.de/fileadmin/fg112/Papers/2019/gawlowicz19_mswim.pdf
2019
Later among the works it cites.
Z. Zhang, S. Zohren, and S. Roberts, “Deep reinforcement learning for trading,” The Journal of Financial Data Science , vol. 2, no. 2, pp. 25–40, 2020
2020
Closest in time.
A. Qayyum, M. Usama, J. Qadir, and A. Al-Fuqaha, “Securing connected & autonomous vehicles: Challenges posed by adversarial machine learning and the way forward,” IEEE Communications Surveys & Tutorials , vol. 22, no. 2, pp. 998–1026, 2020
2020
Closest in time.
J. Sun, T. Zhang, X. Xie, L. Ma, Y. Zheng, K. Chen, and Y. Liu, “Stealthy and efficient adversarial attacks against deep reinforcement learning,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 04, 2020, pp. 5883–5891
2020
Closest in time.
L. Hussenot, M. Geist, and O. Pietquin, “Targeted attacks on deep reinforcement learning agents through adversarial observations,” Autonomous Agents and Multi-Agent Systems (AAMAS) , May 2020
2020
Closest in time.
P. P. Chan, Y. Wang, and D. S. Yeung, “Adversarial attack against deep reinforcement learning with static reward impact map,” in Proceedings of the 15th ACM Asia Conference on Computer and Communications Security , 2020, pp. 334–343
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
A. Gleave, M. Dennis, C. Wild, N. Kant, S. Levine, and S. Russell, “Adversarial policies: Attacking deep reinforcement learning,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=HJgEMpVFwB
2020
Closest in time.
C.-H. H. Yang, J. Qi, P.-Y. Chen, Y. Ouyang, I.-T. D. Hung, C.-H. Lee, and X. Ma, “Enhanced adversarial strategically-timed attacks against deep reinforcement learning,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2020, pp. 3407–3411
2020
Closest in time.
K. Panagiota, W. Kacper, S. Jha, and L. Wenchao, “TrojDRL: Trojan attacks on deep reinforcement learning agents.” in Proc. 57th ACM/IEEE Design Automation Conference (DAC), 2020 , Mar. 2020
2020
Closest in time.
A. Rakhsha, G. Radanovic, R. Devidze, X. Zhu, and A. Singla, “Policy teaching via environment poisoning: Training-time adversarial attacks against reinforcement learning,” in International Conference on Machine Learning . PMLR, 2020, pp. 7974–7984
2020
Closest in time.
X. Y. Lee, S. Ghadai, K. L. Tan, C. Hegde, and S. Sarkar, “Spatiotemporally constrained action space attacks on deep reinforcement learning agents,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 04, 2020, pp. 4577–4584
2020
Closest in time.
M. Huai, J. Sun, R. Cai, L. Yao, and A. Zhang, “Malicious attacks against deep reinforcement learning interpretations,” in Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining , 2020, pp. 472–482
2020
Closest in time.
K. L. Tan, Y. Esfandiari, X. Y. Lee, S. Sarkar et al. , “Robustifying reinforcement learning agents via action space adversarial training,” in 2020 American control conference (ACC) . IEEE, 2020, pp. 3959–3964
2020
Closest in time.
2020
Closest in time.
J. Wang, Y. Liu, and B. Li, “Reinforcement learning with perturbed rewards,” Thirty-forth AAAI Conference on Artificial Intelligence , Feb. 2020
2020
Closest in time.
X. Wang, S. Nair, and M. Althoff, “Falsification-based robust adversarial reinforcement learning,” in 2020 19th IEEE International Conference on Machine Learning and Applications (ICMLA) . IEEE, 2020, pp. 205–212
2020
Closest in time.
B. Lütjens, M. Everett, and J. P. How, “Certified adversarial robustness for deep reinforcement learning,” in Conference on Robot Learning , 2020, pp. 1328–1337
2020
Closest in time.
H. Zhang, H. Chen, C. Xiao, B. Li, M. Liu, D. Boning, and C.-J. Hsieh, “Robust deep reinforcement learning against adversarial perturbations on state observations,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. F. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 21 024–21 037. [Online]. Available: https://proceedings.neurips.cc/paper/2020/file/f0eb6568ea114ba6e293f903c34d7488-Paper.pdf
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
E. Puiutta and E. M. Veith, “Explainable reinforcement learning: A survey,” in International Cross-Domain Conference for Machine Learning and Knowledge Extraction . Springer, 2020, pp. 77–95
2020
Closest in time.
B. R. Kiran, I. Sobh, V. Talpaert, P. Mannion, A. A. Al Sallab, S. Yogamani, and P. Pérez, “Deep reinforcement learning for autonomous driving: A survey,” IEEE Transactions on Intelligent Transportation Systems , 2021
2021
Closest in time.
OpenAI, 2018, [Accessed 21 Jan. 2021.]. [Online]. Available: https://spinningup.openai.com/en/latest/spinningup/rl_intro2.html
2021
Closest in time.
M. Usama, R. Mitra, I. Ilahi, J. Qadir, and M. Marina, “Examining machine learning for 5G and beyond through an adversarial lens,” IEEE Internet Computing , 1 2021
2021
Closest in time.
K. Chen, S. Guo, T. Zhang, X. Xie, and Y. Liu, “Stealing deep reinforcement learning models for fun and profit,” in Proceedings of the 2021 ACM Asia Conference on Computer and Communications Security , 2021, pp. 307–319
2021
Closest in time.
X. Y. Lee, Y. Esfandiari, K. L. Tan, and S. Sarkar, “Query-based targeted action-space adversarial policies on deep reinforcement learning agents,” in Proceedings of the ACM/IEEE 12th International Conference on Cyber-Physical Systems , 2021, pp. 87–97
2021
Closest in time.
2021
Closest in time.
ReAgent, “ReAgent”,” [Accessed Jan. 21, 2021.]. [Online]. Available: https://reagent.ai/
2021
Closest in time.
A. Heuillet, F. Couthouis, and N. Díaz-Rodríguez, “Explainability in deep reinforcement learning,” Knowledge-Based Systems , vol. 214, p. 106685, 2021
2021
Closest in time.
A. Pattanaik, Z. Tang, S. Liu, G. Bommannan, and G. Chowdhary, “Robust deep reinforcement learning with adversarial attacks,” in Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems . International Foundation for Autonomous Agents and Multiagent Systems, Jul. 2018, pp. 2040–2042
2042
Closest in time.