Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) is one of the most popular algorithms to realize an autonomous driving (AD) system.
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra, “Planning and acting in partially observable stochastic domains,” Artificial intelligence , vol. 101, no. 1-2, pp. 99–134, 1998
1998
Earlier work this paper cites.
T. Ayres, L. Li, D. Schleuning, and D. Young, “Preferred time-headway of highway drivers,” in IEEE ITSC , 2001, pp. 826–829
2001
Earlier work this paper cites.
S. Panwai and H. Dia, “Comparative evaluation of microscopic car-following behavior,” IEEE TITS , vol. 6, no. 3, pp. 314–325, 2005
2005
Earlier work this paper cites.
A. Kesting, M. Treiber, and D. Helbing, “General lane-changing model mobil for car-following models,” Transportation Research Record , vol. 1999, no. 1, pp. 86–94, 2007
2007
Earlier work this paper cites.
C. Sarraute, O. Buffet, and J. Hoffmann, “POMDPs make better hackers: Accounting for uncertainty in penetration testing,” in AAAI , 2012
2012
Earlier work this paper cites.
J. Stallkamp, M. Schlipsing, J. Salmen, and C. Igel, “Man vs. computer: Benchmarking machine learning algorithms for traffic sign recognition,” Neural networks , vol. 32, pp. 323–332, 2012
2012
Earlier work this paper cites.
M. Hausknecht and P. Stone, “Deep recurrent q-learning for partially observable MDPs,” in AAAI , 2015
2015
Earlier work this paper cites.
J. Kong, M. Pfeiffer, G. Schildbach, and F. Borrelli, “Kinematic and dynamic vehicle models for autonomous driving control design,” in IEEE Intelligent Vehicles Symposium (IV) . IEEE, 2015, pp. 1094–1099
2015
Earlier work this paper cites.
H. Bai, S. Cai, N. Ye, D. Hsu, and W. S. Lee, “Intention-aware online pomdp planning for autonomous driving in a crowd,” in ICRA . IEEE, 2015, pp. 454–460
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
H. Mao, R. Netravali, and M. Alizadeh, “Neural adaptive video streaming with pensieve,” in ACM SIGCOMM , 2017, pp. 197–210
2017
Earlier work this paper cites.
Y. F. Chen, M. Everett, M. Liu, and J. P. How, “Socially aware motion planning with deep reinforcement learning,” in IEEE/RSJ IROS , 2017, pp. 1343–1350
2017
Earlier work this paper cites.
M. Igl, L. Zintgraf, T. A. Le, F. Wood, and S. Whiteson, “Deep variational reinforcement learning for POMDPs,” in ICML , 2018, pp. 2117–2126
2018
Earlier work this paper cites.
B. Tran, J. Li, and A. Madry, “Spectral signatures in backdoor attacks,” NeurIPS , vol. 31, 2018
2018
Earlier work this paper cites.
E. Leurent, “An environment for autonomous driving decision-making,” https://github.com/eleurent/highway-env , 2018
2018
Earlier work this paper cites.
Red and B. Teams, “Adversarial robustness toolbox (art) - python library for machine learning security - evasion, poisoning, extraction, inference,” https://github.com/Trusted-AI/adversarial-robustness-toolbox , 2018
2018
Earlier work this paper cites.
C. Wang, J. Wang, Y. Shen, and X. Zhang, “Autonomous navigation of uavs in large-scale complex environments: A deep reinforcement learning approach,” IEEE TVT , vol. 68, no. 3, pp. 2124–2136, 2019
2019
Cited alongside, same era.
Z. Pan, Y. Liang, W. Wang, Y. Yu, Y. Zheng, and J. Zhang, “Urban traffic prediction from spatio-temporal data using deep meta learning,” in ACM KDD , 2019, pp. 1720–1730
2019
Cited alongside, same era.
Y. Chen, C. Dong, P. Palanisamy, P. Mudalige, K. Muelling, and J. M. Dolan, “Attention-based hierarchical deep reinforcement learning for lane change behaviors in autonomous driving,” in CVPR , 2019
2019
Cited alongside, same era.
E. Leurent and J. Mercat, “Social attention for autonomous decision-making in dense traffic,” in Machine Learning for Autonomous Driving Workshop at NeurIPS , 2019
2019
Cited alongside, same era.
C. Ashcraft and K. Karra, “Poisoning deep reinforcement learning agents with in-distribution triggers,” ICLR Workshop , 2021
2021
Later among the works it cites.
H. Seong, C. Jung, S. Lee, and D. H. Shim, “Learning to drive at unsignalized intersections using attention-based deep reinforcement learning,” in IEEE ITSC , 2021, pp. 559–566
2021
Later among the works it cites.
X. Ma, J. Li, M. J. Kochenderfer, D. Isele, and K. Fujimura, “Reinforcement learning for autonomous driving with latent state inference and spatial-temporal relationships,” in ICRA . IEEE, 2021, pp. 6064–6071
2021
Later among the works it cites.
L. Wang, Z. Javed, X. Wu, W. Guo, X. Xing, and D. Song, “Backdoorl: Backdoor attack against competitive reinforcement learning,” in IJCAI , 2021
2021
Later among the works it cites.
S. Cheng, Y. Liu, S. Ma, and X. Zhang, “Deep feature space trojan attack of neural networks by controlled detoxification,” in AAAI , vol. 35, no. 2, 2021, pp. 1148–1156
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Y. Gao, C. Xu, D. Wang, S. Chen, D. C. Ranasinghe, and S. Nepal, “Strip: A defence against trojan attacks on deep neural networks,” in ACM CCS , 2019, pp. 113–125
2019
Cited alongside, same era.
B. Wang, Y. Yao, S. Shan, H. Li, B. Viswanath, H. Zheng, and B. Y. Zhao, “Neural cleanse: Identifying and mitigating backdoor attacks in neural networks,” in IEEE S&P , 2019, pp. 707–723
2019
Cited alongside, same era.
H. Chen, C. Fu, J. Zhao, and F. Koushanfar, “Deepinspect: A black-box trojan detection and mitigation framework for deep neural networks.” in IJCAI , 2019
2019
Cited alongside, same era.
B. Chen, W. Carvalho, N. Baracaldo, H. Ludwig, B. Edwards, T. Lee, I. Molloy, and B. Srivastava, “Detecting backdoor attacks on deep neural networks by activation clustering,” in Artificial Intelligence Safety Workshop at AAAI , 2019
2019
Cited alongside, same era.
K. Muhammad, A. Ullah, J. Lloret, J. Del Ser, and V. H. C. de Albuquerque, “Deep learning for safe autonomous driving: Current challenges and future directions,” IEEE TITS , vol. 22, no. 7, pp. 4316–4336, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
P. Kiourti, K. Wardega, S. Jha, and W. Li, “TrojDRL: evaluation of backdoor attacks on deep reinforcement learning,” in ACM/IEEE DAC , 2020, pp. 1–6
2020
Cited alongside, same era.
2021
Later among the works it cites.
K. Doan, Y. Lao, W. Zhao, and P. Li, “Lira: Learnable, imperceptible and robust backdoor attacks,” in ICCV , 2021, pp. 11 966–11 976
2021
Later among the works it cites.
T. A. Nguyen and A. T. Tran, “Wanet-imperceptible warping-based backdoor attack,” in ICLR , 2021
2021
Later among the works it cites.
X. Gong, Y. Chen, Q. Wang, H. Huang, L. Meng, C. Shen, and Q. Zhang, “Defense-resistant backdoor attacks against deep neural networks in outsourced cloud environment,” IEEE JSAC , vol. 39, no. 8, pp. 2617–2631, 2021
2021
Later among the works it cites.
B. R. Kiran, I. Sobh, V. Talpaert, P. Mannion, A. A. Al Sallab, S. Yogamani, and P. Pérez, “Deep reinforcement learning for autonomous driving: A survey,” IEEE TITS , vol. 23, no. 6, pp. 4909–4926, 2022
2022
Closest in time.
S. Aradi, “Survey of deep reinforcement learning for motion planning of autonomous vehicles,” IEEE TITS , vol. 23, no. 2, pp. 740–759, 2022
2022
Closest in time.
Y. Yu, J. Liu, and J. Fang, “Online microservice orchestration for iot via multi-objective deep reinforcement learning,” IEEE IoTJ , 2022
2022
Closest in time.
S. Mo, X. Pei, and C. Wu, “Safe reinforcement learning for autonomous vehicle using monte carlo tree search,” IEEE TITS , vol. 23, no. 7, pp. 6766–6773, 2022
2022
Closest in time.
Y. Li, Y. Jiang, Z. Li, and S.-T. Xia, “Backdoor learning: A survey,” IEEE TNNLS , pp. 1–18, 2022
2022
Closest in time.
Y. Yu, J. Liu, S. Li, K. Huang, and X. Feng, “A temporal-pattern backdoor attack to deep reinforcement learning,” in IEEE GLOBECOM , 2022
2022
Closest in time.
2022
Closest in time.
T. Ni, B. Eysenbach, and R. Salakhutdinov, “Recurrent model-free rl can be a strong baseline for many pomdps,” in ICML , 2022, pp. 16 691–16 723
2022
Closest in time.