Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) has made significant achievements in many real-world applications.
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra, “Planning and acting in partially observable stochastic domains,” Artificial intelligence , vol. 101, no. 1-2, pp. 99–134, 1998
1998
Earlier work this paper cites.
C. Sarraute, O. Buffet, and J. Hoffmann, “POMDPs make better hackers: Accounting for uncertainty in penetration testing,” in AAAI , 2012
2012
Earlier work this paper cites.
M. Hausknecht and P. Stone, “Deep recurrent q-learning for partially observable mdps,” in AAAI , 2015
2015
Earlier work this paper cites.
H. Mao, R. Netravali, and M. Alizadeh, “Neural adaptive video streaming with pensieve,” in ACM SIGCOMM , 2017, pp. 197–210
2017
Earlier work this paper cites.
M. Igl, L. Zintgraf, T. A. Le, F. Wood, and S. Whiteson, “Deep variational reinforcement learning for POMDPs,” in ICML , 2018, pp. 2117–2126
2018
Earlier work this paper cites.
B. Tran, J. Li, and A. Madry, “Spectral signatures in backdoor attacks,” NeurIPS , vol. 31, 2018
2018
Earlier work this paper cites.
Y. Wei, L. Pan, S. Liu, L. Wu, and X. Meng, “DRL-scheduling: An intelligent QoS-aware job scheduling framework for applications in clouds,” IEEE Access , vol. 6, pp. 55 112–55 125, 2018
2018
Earlier work this paper cites.
N. C. Luong, D. T. Hoang, S. Gong, D. Niyato, P. Wang, Y.-C. Liang, and D. I. Kim, “Applications of deep reinforcement learning in communications and networking: A survey,” IEEE Comm. Surveys & Tutorials , vol. 21, no. 4, pp. 3133–3174, 2019
2019
Earlier work this paper cites.
2019
Cited alongside, same era.
Y. Gao, C. Xu, D. Wang, S. Chen, D. C. Ranasinghe, and S. Nepal, “Strip: A defence against trojan attacks on deep neural networks,” in ACM CCS , 2019, pp. 113–125
2019
Cited alongside, same era.
B. Wang, Y. Yao, S. Shan, H. Li, B. Viswanath, H. Zheng, and B. Y. Zhao, “Neural cleanse: Identifying and mitigating backdoor attacks in neural networks,” in IEEE Symp. S&P , 2019, pp. 707–723
2019
Cited alongside, same era.
2020
Cited alongside, same era.
J. Zhao, F. Huang, J. Lv, Y. Duan, Z. Qin, G. Li, and G. Tian, “Do RNN and LSTM have long memory?” in ICML , 2020, pp. 11 365–11 375
2020
Later among the works it cites.
C. Ashcraft and K. Karra, “Poisoning deep reinforcement learning agents with in-distribution triggers,” ICLR Workshop , 2021
2021
Later among the works it cites.
S. Mo, X. Pei, and C. Wu, “Safe reinforcement learning for autonomous vehicle using monte carlo tree search,” IEEE TITS , pp. 1–8, 2021
2021
Later among the works it cites.
Y. Wang, E. Sarkar, W. Li, M. Maniatakos, and S. E. Jabari, “Stop-and-go: Exploring backdoor attacks on deep reinforcement learning-based traffic congestion control systems,” IEEE TIFS , vol. 16, pp. 4772–4787, 2021
2021
Later among the works it cites.
L. Wang, Z. Javed, X. Wu, W. Guo, X. Xing, and D. Song, “Backdoorl: Backdoor attack against competitive reinforcement learning,” in IJCAI , 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
N. Yang, H. Zhang, and R. Berry, “Partially observable multi-agent deep reinforcement learning for cognitive resource management,” in GLOBECOM . IEEE, 2020, pp. 1–6
2020
Cited alongside, same era.
Y. Zhan, S. Guo, P. Li, and J. Zhang, “A deep reinforcement learning based offloading game in edge computing,” IEEE Trans. Comput. , vol. 69, no. 6, pp. 883–893, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2021
Later among the works it cites.
X. Gong, Y. Chen, Q. Wang, H. Huang, L. Meng, C. Shen, and Q. Zhang, “Defense-resistant backdoor attacks against deep neural networks in outsourced cloud environment,” IEEE JSAC , vol. 39, no. 8, pp. 2617–2631, 2021
2021
Later among the works it cites.
Y. Huang, L. Cheng, L. Xue, C. Liu, Y. Li, J. Li, and T. Ward, “Deep adversarial imitation reinforcement learning for QoS-aware cloud job scheduling,” IEEE Systems Journal , 2021
2021
Later among the works it cites.