Fetching the paper…
Reading the bibliography…
The scale of Internet-connected systems has increased considerably, and these systems are being exposed to cyber attacks more than ever.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, … and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International Conference on Machine Learning
1937
Earlier work this paper cites.
R. S. Sutton, “Learning to predict by the methods of temporal differences,” Machine Learning
1988
Earlier work this paper cites.
R. S. Sutton, “Integrated architecture for learning, planning, and reacting based on approximating dynamic programming,” in The 7th International Conference on Machine Learning
1990
Earlier work this paper cites.
C. J. Watkins, and P. Dayan, “Q-learning,” Machine Learning
1992
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine Learning
1992
Earlier work this paper cites.
M. L. Littman, “Markov games as a framework for multiagent reinforcement learning,” in The 11th International Conference on Machine Learning
1994
Earlier work this paper cites.
S. Hochreiter, and J. Schmidhuber, “Long short-term memory,” Neural Computation
1997
Earlier work this paper cites.
R. S. Sutton, A. G. Barto, “Introduction to Reinforcement Learning,” MIT Press Cambridge, MA, USA, 1998
1998
Earlier work this paper cites.
R. S. Sutton, D. A. McAllester, S. P. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” in Advances in Neural Information Processing Systems
2000
Earlier work this paper cites.
M. Bowling, and M. Veloso, “Multiagent learning using a variable learning rate,” Artificial Intelligence
2002
Earlier work this paper cites.
Z. Wang, T. Schaul, M. Hessel, H. Hasselt, M. Lanctot, and N. Freitas, “Dueling network architectures for deep reinforcement learning,” in International Conference on Machine Learning
2003
Earlier work this paper cites.
X. Xu, and T. Xie, “A reinforcement learning approach for host-based intrusion detection using sequences of system calls,” in International Conference on Intelligent Computing
2005
Earlier work this paper cites.
D. K. Yau, J. C. Lui, F. Liang, and Y. Yam, “Defending against distributed denial-of-service attacks with max-min fair server-centric router throttles,” IEEE/ACM Transactions on Networking
2005
Earlier work this paper cites.
A. Detwarasiti, and R. D. Shachter, “Influence diagrams for team decision analysis,” Decision Analysis
2005
Earlier work this paper cites.
X. Xu, “A sparse kernel-based least-squares temporal difference algorithm for reinforcement learning,” in International Conference on Natural Computation
2006
Earlier work this paper cites.
X. Xu, and Y. Luo, “A kernel-based reinforcement learning approach to dynamic behavior modeling of intrusion detection,” in International Symposium on Neural Networks
2007
Earlier work this paper cites.
A. Herrero, and E. Corchado, “Multiagent systems for network intrusion detection: A review,” in Computational Intelligence in Security for Information Systems
2009
Earlier work this paper cites.
A. Mpitziopoulos, D. Gavalas, C. Konstantopoulos, and G. Pantziou, “A survey on jamming attacks and countermeasures in WSNs,” IEEE Communications Surveys and Tutorials
2009
Earlier work this paper cites.
X. Xu, “Sequential anomaly detection based on temporal-difference learning: Principles, models and case studies,” Applied Soft Computing
2010
Earlier work this paper cites.
S. Roy, C. Ellis, S. Shiva, D. Dasgupta, V. Shandilya, and Q. Wu, “A survey of game theory as applied to network security,” in 43rd Hawaii International Conference on System Sciences
2010
Earlier work this paper cites.
S. Shiva, S. Roy, and D. Dasgupta, “Game theory for cyber security,” in The Sixth Annual Workshop on Cyber Security and Information Intelligence Research
2010
Earlier work this paper cites.
K. Zeng, K. Govindan, and P. Mohapatra, “Non-cryptographic authentication and identification in wireless networks,” IEEE Wireless Communications
2010
Earlier work this paper cites.
A. Sharma, Z. Kalbarczyk, J. Barlow, and R. Iyer, “Analysis of security data from a large computing organization,” in Dependable Systems and Networks (DSN), IEEE/IFIP 41st International Conference on
2011
Earlier work this paper cites.
B. Wang, Y. Wu, K. R. Liu, and T. C. Clancy, “An anti-jamming stochastic game for cognitive radio networks,” IEEE Journal on Selected Areas in Communications
2011
Earlier work this paper cites.
H. Abbas, and G. Fainekos, “Convergence proofs for simulated annealing falsification of safety properties,” in 2012 50th Annual Allerton Conference on Communication, Control, and Computing (Allerton)
2012
Earlier work this paper cites.
S. Sankaranarayanan, and G. Fainekos, “Falsification of temporal properties of hybrid systems using the cross-entropy method,” in The 15th ACM International Conference on Hybrid Systems: Computation and Control
2012
Earlier work this paper cites.
B. Deokar, and A. Hazarnis, “Intrusion detection system using log files and reinforcement learning,” International Journal of Computer Applications
2012
Earlier work this paper cites.
Y. Wu, B. Wang, K. R. Liu, and T. C. Clancy, “Anti-jamming games in multi-channel cognitive radio networks,” IEEE Journal on Selected Areas in Communications
2012
Earlier work this paper cites.
S. Singh, and A. Trivedi, “Anti-jamming in cognitive radio networks using reinforcement learning algorithms,” in Wireless and Optical Communications Networks (WOCN), Ninth International Conference on
2012
Earlier work this paper cites.
A. Attar, H. Tang, A. V. Vasilakos, F. R. Yu, and V. C. Leung, “A survey of security challenges in cognitive radio networks: Solutions and future research directions,” Proceedings of the IEEE
2012
Earlier work this paper cites.
C. B. Browne, E. Powley, D. Whitehouse, S. M. Lucas, P. I. Cowling, P. Rohlfshagen, … and S. Colton, “A survey of Monte Carlo tree search methods,” IEEE Transactions on Computational Intelligence and AI in Games
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
P. Muñoz, R. Barco, and I. de la Bandera, “Optimization of load balancing using fuzzy Q-learning for next generation wireless networks,” Expert Systems with Applications
2013
Earlier work this paper cites.
S. Shamshirband, N. B. Anuar, M. L. M. Kiah, and A. Patel, “An appraisal and design of a multiagent system based cooperative wireless intrusion detection computational intelligence technique,” Engineering Applications of Artificial Intelligence
2013
Earlier work this paper cites.
Y. Gwon, S. Dastangoo, C. Fossa, and H. T. Kung, “Competing mobile network game: Embracing antijamming and jamming strategies with reinforcement learning,” in IEEE Conference on Communications and Network Security (CNS)
2013
Earlier work this paper cites.
W. G. Conley, and A. J. Miller, “Cognitive jamming game for dynamically countering ad hoc cognitive radio networks,” in MILCOM 2013-2013 IEEE Military Communications Conference
2013
Earlier work this paper cites.
D. Yang, G. Xue, J. Zhang, A. Richa, and X. Fang, “Coping with a smart jammer in wireless networks: A Stackelberg game approach,” IEEE Transactions on Wireless Communications
2013
Earlier work this paper cites.
K. Malialis, “Distributed Reinforcement Learning for Network Intrusion Response,” Doctoral Dissertation, University of York, UK, 2014
2014
Earlier work this paper cites.
R. Bhosale, S. Mahajan, and P. Kulkarni, “Cooperative machine learning for intrusion detection system,” International Journal of Scientific and Engineering Research
2014
Earlier work this paper cites.
S. Shamshirband, A. Patel, N. B. Anuar, M. L. M. Kiah, and A. Abraham, “Cooperative game theoretic approach using fuzzy Q-learning for detecting and preventing intrusions in wireless sensor networks,” Engineering Applications of Artificial Intelligence
2014
Earlier work this paper cites.
K. Dabcevic, A. Betancourt, L. Marcenaro, and C. S. Regazzoni, “A fictitious play-based game-theoretical approach to alleviating jamming attacks for cognitive radios,” in 2014 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2014
Earlier work this paper cites.
M. Zhu, Z. Hu, and P. Liu, “Reinforcement learning algorithms for adaptive cyber defense against Heartbleed,” in The First ACM Workshop on Moving Target Defense
2014
Earlier work this paper cites.
M. H. Ling, K. L. A. Yau, J. Qadir, G. S. Poh, and Q. Ni, “Application of reinforcement learning for security enhancement in cognitive radio networks,” Applied Soft Computing
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, … and S. Petersen, “Human-level control through deep reinforcement learning,” Nature
2015
Earlier work this paper cites.
T. Schaul, J. Quan, I. Antonoglou, and D. Silver, “Prioritized experience replay,” arXiv preprint
2015
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International Conference on Machine Learning
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
L. Wang, M. Torngren, and M. Onori, “Current status and advancement of cyber-physical systems in manufacturing,” Journal of Manufacturing Systems
2015
Earlier work this paper cites.
K. Malialis, and D. Kudenko, “Distributed response to network intrusions using multiagent reinforcement learning,” Engineering Applications of Artificial Intelligence
2015
Earlier work this paper cites.
F. Slimeni, B. Scheers, Z. Chtourou, and V. Le Nir, “Jamming mitigation in cognitive radio networks using a modified Q-learning algorithm,” in 2015 International Conference on Military Communications and Information Systems (ICMCIS)
2015
Earlier work this paper cites.
L. Xiao, Y. Li, J. Liu, and Y. Zhao, “Power control with reinforcement learning in cooperative cognitive radio networks against jamming,” The Journal of Supercomputing
2015
Earlier work this paper cites.
L. Xiao, Y. Li, G. Liu, Q. Li, and W. Zhuang, “Spoofing detection with reinforcement learning in wireless networks,” in Global Communications Conference (GLOBECOM)
2015
Earlier work this paper cites.
Y. Li, J. Liu, Q. Li, and L. Xiao, “Mobile cloud offloading for malware detections with learning,” in IEEE Conference on Computer Communications Workshops
2015
Earlier work this paper cites.
M. A. Salahuddin, A. Al-Fuqaha, and M. Guizani, “Software-defined networking for RSU clouds in support of the internet of vehicles,” IEEE Internet of Things Journal
2015
Earlier work this paper cites.
R. Huang, X. Chu, J. Zhang, and Y. H. Hu, “Energy-efficient monitoring in software defined wireless sensor networks using reinforcement learning: A prototype,” International Journal of Distributed Sensor Networks
2015
Earlier work this paper cites.
B. Lantz, and B. O’Connor, “A mininet-based virtual testbed for distributed SDN development,” in ACM SIGCOMM Computer Communication Review
2015
Earlier work this paper cites.
J. Wang, M. Zhao, Q. Zeng, D. Wu, and P. Liu, “Risk assessment of buffer ”Heartbleed” over-read vulnerabilities,” in 45th Annual IEEE/IFIP International Conference on Dependable Systems and Networks
2015
Earlier work this paper cites.
J. Oh, X. Guo, H. Lee, R. L. Lewis, and S. Singh, “Action-conditional video prediction using deep networks in atari games,” in Advances in Neural Information Processing Systems
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. Botta, W. De Donato, V. Persico, and A. Pescapé, “Integration of cloud computing and Internet of Things: a survey,” Future Generation Computer Systems
2016
Earlier work this paper cites.
A. V. Dastjerdi, and R. Buyya, “Fog computing: Helping the Internet of Things realize its potential,” Computer
2016
Earlier work this paper cites.
A. L. Buczak, and E. Guven, “A survey of data mining and machine learning methods for cyber security intrusion detection,” IEEE Communications Surveys and Tutorials
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, … and S. Dieleman, “Mastering the game of Go with deep neural networks and tree search,” Nature
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
W. Haider, G. Creech, Y. Xie, and J. Hu, “Windows based data sets for evaluation of robustness of host based intrusion detection systems (IDS) to zero-day and stealth attacks,” Future Internet
2016
Earlier work this paper cites.
M. Nobakht, V. Sivaraman, and R. Boreli, “A host-based intrusion detection and mitigation framework for smart home IoT using OpenFlow,” in 11th International Conference on Availability, Reliability and Security (ARES)
2016
Earlier work this paper cites.
2016
Cited alongside, same era.
K. Ramachandran, and Z. Stefanova, “Dynamic game theories in cyber security,” in International Conference of Dynamic Systems and Applications
2016
Cited alongside, same era.
Y. Wang, Y. Wang, J. Liu, Z. Huang, and P. Xie, “A survey of game theoretic methods for cyber security,” in IEEE First International Conference on Data Science in Cyberspace (DSC)
2016
Cited alongside, same era.
S. Machuzak, and S. K. Jayaweera, “Reinforcement learning based anti-jamming with wideband autonomous cognitive radios,” in 2016 IEEE/CIC International Conference on Communications in China (ICCC)
2016
Cited alongside, same era.
S. Varshney, and R. Kuma, “Variants of LEACH routing protocol in WSN: A comparative analysis,” in The 8th International Conference on Cloud Computing, Data Science and Engineering (Confluence)
2018
Later among the works it cites.
Q. Zhu, and S. Rass, “Game theory meets network security: A tutorial,” in The 2018 ACM SIGSAC Conference on Computer and Communications Security
2018
Later among the works it cites.
S. Hu, D. Yue, X. Xie, X. Chen, and X. Yin, “Resilient event-triggered controller synthesis of networked control systems under periodic DoS jamming attacks,” IEEE Transactions on Cybernetics
2018
Later among the works it cites.
L. Xiao, X. Lu, D. Xu, Y. Tang, L. Wang, and W. Zhuang, “UAV relay in VANETs against smart jamming with reinforcement learning,” IEEE Transactions on Vehicular Technology
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Chen, and X. Wen, “Perceptual spectrum waterfall of pattern shape recognition algorithm,” in The 18th International Conference on Advanced Communication Technology (ICACT)
2016
Cited alongside, same era.
L. Xiao, Y. Li, G. Han, G. Liu, and W. Zhuang, “PHY-layer spoofing detection with reinforcement learning in wireless networks,” IEEE Transactions on Vehicular Technology
2016
Cited alongside, same era.
S. Kim, J. Son, A. Talukder, and C. S. Hong, “Congestion prevention mechanism based on Q-leaning for efficient routing in SDN,” in International Conference on Information Networking (ICOIN)
2016
Cited alongside, same era.
S. C. Lin, I. F. Akyildiz, P. Wang, and M. Luo, “QoS-aware adaptive routing in multi-layer hierarchical software defined networks: A reinforcement learning approach,” in 2016 IEEE International Conference on Services Computing (SCC)
2016
Cited alongside, same era.
K. Chung, C. A. Kamhoua, K. A. Kwiat, Z. T. Kalbarczyk, and R. K. Iyer, “Game theory with learning for cyber security monitoring,” in IEEE 17th International Symposium on High Assurance Systems Engineering (HASE)
2016
Cited alongside, same era.
A. Tamar, Y. Wu, G. Thomas, S. Levine, and P. Abbeel, “Value iteration networks,” in Advances in Neural Information Processing Systems
2016
Cited alongside, same era.
I. Kakalou, K. E. Psannis, P. Krawiec, and R. Badea, “Cognitive radio network and network service chaining toward 5G: challenges and requirements,” IEEE Communications Magazine
2017
Cited alongside, same era.
N. Milosevic, A. Dehghantanha, and K. K. R. Choo, “Machine learning aided Android malware classification,” Computers and Electrical Engineering
2017
Cited alongside, same era.
2018
Later among the works it cites.
X. Liu, Y. Xu, L. Jia, Q. Wu, and A. Anpalagan, “Anti-jamming communications using spectrum waterfall: A deep reinforcement learning approach,” IEEE Communications Letters
2018
Later among the works it cites.
X. Sun, J. Dai, P. Liu, A. Singhal, and J. Yen, “Using Bayesian networks for probabilistic identification of zero-day attack paths,” IEEE Transactions on Information Forensics and Security
2018
Later among the works it cites.
Y. Han, B. I. Rubinstein, T. Abraham, T. Alpcan, O. De Vel, S. Erfani, … and P. Montague, “Reinforcement learning for autonomous defence in software-defined networking,” in International Conference on Decision and Game Theory for Security
2018
Later among the works it cites.
B. Luo, Y. Yang, C. Zhang, Y. Wang, and B. Zhang, “A survey of code reuse attack and defense,” in International Conference on Intelligent and Interactive Systems and Applications
2018
Later among the works it cites.
A. Nagabandi, G. Kahn, R. S. Fearing, and S. Levine, “Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA)
2018
Later among the works it cites.
V. Duddu, “A survey of adversarial machine learning in cyber warfare,” Defence Science Journal
2018
Later among the works it cites.
Y. Liu, M. Dong, K. Ota, J. Li, and J. Wu, “Deep reinforcement learning based smart mitigation of DDoS flooding in software-defined networks,” in IEEE 23rd International Workshop on Computer Aided Modeling and Design of Communication Links and Networks (CAMAD)
2018
Later among the works it cites.
Y. Xu, G. Ren, J. Chen, X. Zhang, L. Jia, and L. Kong, “Interference-aware cooperative anti-jamming distributed channel selection in UAV communication networks,” Applied Sciences
2018
Later among the works it cites.
B. Geluvaraj, P. M. Satwik, and T. A. Kumar, “The future of cybersecurity: major role of artificial intelligence, machine learning, and deep learning in cyberspace,” in International Conference on Computer Networks and Communication Technologies
2019
Closest in time.
D. S. Berman, A. L. Buczak, J. S. Chavis, and C. L. Corbett, “A survey of deep learning methods for cyber security,” Information
2019
Closest in time.
M. Wu, Z. Song, and Y. B. Moon, “Detecting cyber-physical attacks in CyberManufacturing systems with machine learning methods,” Journal of Intelligent Manufacturing
2019
Closest in time.
Y. Wang, Z. Ye, P. Wan, and J. Zhao, “A survey of dynamic spectrum allocation based on reinforcement learning algorithms in cognitive radio networks,” Artificial Intelligence Review
2019
Closest in time.
OpenAI, “OpenAI Five,” [Online]. Available: https://openai.com/five/, 2019, March 1
2019
Closest in time.
T. T. Nguyen, N. D. Nguyen, F. Bello, and S. Nahavandi, “A new tensioning method using deep reinforcement learning for surgical pattern cutting,” in 2019 IEEE International Conference on Industrial Technology (ICIT)
2019
Closest in time.
N. D. Nguyen, T. Nguyen, S. Nahavandi, A. Bhatti, and G. Guest, “Manipulating soft tissues by deep reinforcement learning for autonomous robotic surgery,” in 2019 IEEE International Systems Conference (SysCon)
2019
Closest in time.
N. C. Luong, D. T. Hoang, S. Gong, D. Niyato, P. Wang, Y. C. Liang, and D. I. Kim, “Applications of deep reinforcement learning in communications and networking: A survey,” IEEE Communications Surveys and Tutorials
2019
Closest in time.
Y. Dai, D. Xu, S. Maharjan, Z. Chen, Q. He, and Y. Zhang, “Blockchain and deep reinforcement learning empowered intelligent 5G beyond,” IEEE Network
2019
Closest in time.
Z. Ni, and S. Paul, “A multistage game in smart grid security: A reinforcement learning solution,” IEEE Transactions on Neural Networks and Learning Systems
2019
Closest in time.
C. Li, and M. Qiu, Reinforcement Learning for Cyber-Physical Systems: with Cybersecurity Case Studies
2019
Closest in time.
X. Wang, R. Jiang, L. Li, Y. L. Lin, and F. Y. Wang, “Long memory is important: A test study on deep-learning based car-following model,” Physica A: Statistical Mechanics and its Applications
2019
Closest in time.
S. Dey, Q. Ye, and S. Sampalli, “A machine learning based intrusion detection scheme for data fusion in mobile clouds involving heterogeneous client networks,” Information Fusion
2019
Closest in time.
D. Papamartzivanos, F. G. Mármol, and G. Kambourakis, “Introducing deep learning self-adaptive misuse network intrusion detection systems,” IEEE Access
2019
Closest in time.
G. Caminero, M. Lopez-Martin, and B. Carro, “Adversarial environment reinforcement learning algorithm for intrusion detection,” Computer Networks
2019
Closest in time.
H. Boche, and C. Deppe, “Secure identification under passive eavesdroppers and active jamming attacks,” IEEE Transactions on Information Forensics and Security
2019
Closest in time.
M.D. Felice, L. Bedogni, L. Bononi, “Reinforcement learning-based spectrum management for cognitive radio networks: A literature review and case study,” in Handbook of Cognitive Radio
2019
Closest in time.
Y. Afek, A. Bremler-Barr, and S. L. Feibish, “Zero-day signature extraction for high-volume attacks,” IEEE/ACM Transactions on Networking
2019
Closest in time.
2019
Closest in time.
M. Giles, “Five emerging cyber-threats to worry about in 2019,” MIT Technology Review
2019
Closest in time.
T. Chen, J. Liu, Y. Xiang, W. Niu, E. Tong, and Z. Han, “Adversarial attack and defense in reinforcement learning-from AI security view,” Cybersecurity
2019
Closest in time.
T. Nguyen, N. D. Nguyen, and S. Nahavandi, “Multi-agent deep reinforcement learning with human strategies,” in 2019 IEEE International Conference on Industrial Technology (ICIT)
2019
Closest in time.
F. Yao, and L. Jia, “A collaborative multiagent reinforcement learning anti-jamming algorithm in wireless networks,” IEEE Wireless Communications Letters
2019
Closest in time.
Y. Li, X. Wang, D. Liu, Q. Guo, X. Liu, J. Zhang, and Y. Xu, “On the performance of deep reinforcement learning-based anti-jamming method confronting intelligent jammer,” Applied Sciences
2019
Closest in time.
2019
Closest in time.
Y. Huang, S. Li, C. Li, Y. T. Hou, and W. Lou, “A deep reinforcement learning-based approach to dynamic eMBB/URLLC multiplexing in 5G NR,” IEEE Internet of Things Journal
2020
Closest in time.
P. Wang, L. T. Yang, X. Nie, Z. Ren, J. Li, and L. Kuang, “Data-driven software defined network attack detection: state-of-the-art and perspectives,” Information Sciences
2020
Closest in time.
O. Krestinskaya, A. P. James, and L. O. Chua, “Neuromemristive circuits for edge computing: a review,” IEEE Transactions on Neural Networks and Learning Systems
2020
Closest in time.
S. Paul, Z. Ni, and C. Mu, “A learning-based solution for an adversarial repeated game in cyber-physical power systems,” IEEE Transactions on Neural Networks and Learning Systems
2020
Closest in time.
Z. Sui, Z. Pu, J. Yi, and S. Wu, “Formation control with collision avoidance through deep reinforcement learning using model-guided demonstration,” IEEE Transactions on Neural Networks and Learning Systems
2020
Closest in time.
A. Tsantekidis, N. Passalis, A. S. Toufa, K. Saitas-Zarkias, S. Chairistanidis, and A. Tefas, “Price trailing for financial trading using deep reinforcement learning,” IEEE Transactions on Neural Networks and Learning Systems
2020
Closest in time.
X. Wang, Y. Gu, Y. Cheng, A. Liu, and C. P. Chen, “Approximate policy-based accelerated deep reinforcement learning,” IEEE Transactions on Neural Networks and Learning Systems
2020
Closest in time.
X. Lu, L. Xiao, T. Xu, Y. Zhao, Y. Tang, and W. Zhuang, “Reinforcement learning based PHY authentication for VANETs,” IEEE Transactions on Vehicular Technology
2020
Closest in time.
M. Alauthman, N. Aslam, M. Al-Kasassbeh, S. Khan, A. Al-Qerem, and K. K. R. Choo, “An efficient reinforcement learning-based botnet detection approach,” Journal of Network and Computer Applications
2020
Closest in time.
Y. Keneshloo, T. Shi, N. Ramakrishnan, and C. K. Reddy, “Deep reinforcement learning for sequence-to-sequence models,” IEEE Transactions on Neural Networks and Learning Systems
2020
Closest in time.
R. Shafin, H. Chen, Y. H. Nam, S. Hur, J. Park, J. Zhang, … and L. Liu, “Self-tuning sectorization: deep reinforcement learning meets broadcast beam optimization,” IEEE Transactions on Wireless Communications
2020
Closest in time.
A. S. Leong, A. Ramaswamy, D. E. Quevedo, H. Karl, and L. Shi, “Deep reinforcement learning for wireless sensor scheduling in cyber–physical systems,” Automatica
2020
Closest in time.
OpenAI Gym Toolkit Documentation. Classic control: control theory problems from the classic RL literature. Retrieved December 14, 2020, from: https://gym.openai.com/envs/#classic_control
2020
Closest in time.
OpenAI Gym Toolkit Documentation. Box2D: continuous control tasks in the Box2D simulator. Retrieved December 14, 2020, from: https://gym.openai.com/envs/#box2d
2020
Closest in time.
A. Ferdowsi, A. Eldosouky, and W. Saad, “Interdependence-aware game-theoretic framework for secure intelligent transportation systems,” IEEE Internet of Things Journal
2020
Closest in time.
I. Rasheed, F. Hu, and L. Zhang, “Deep reinforcement learning approach for autonomous vehicle systems for maintaining security and safety using LSTM-GAN,” Vehicular Communications
2020
Closest in time.
M. M. Hassan, A. Gumaei, A. Alsanad, M. Alrubaian, and G. Fortino, “A hybrid deep learning model for efficient intrusion detection in big data environment,” Information Sciences
2020
Closest in time.
M. Lopez-Martin, B. Carro, and A. Sanchez-Esguevillas, “Application of deep reinforcement learning to intrusion detection for supervised problems,” Expert Systems with Applications
2020
Closest in time.
I. A. Saeed, A. Selamat, M. F. Rohani, O. Krejcar, and J. A. Chaudhry, “A systematic state-of-the-art analysis of multiagent intrusion detection,” IEEE Access
2020
Closest in time.
F. Wei, Z. Wan, and H. He, “Cyber-attack recovery strategy for smart grid based on deep reinforcement learning,” IEEE Transactions on Smart Grid
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
L. Tong, A. Laszka, C. Yan, N. Zhang, and Y. Vorobeychik, “Finding needles in a moving haystack: prioritizing alerts with adversarial reinforcement learning,” in Proceedings of the AAAI Conference on Artificial Intelligence
2020
Closest in time.
2020
Closest in time.
T. T. Nguyen, N. D. Nguyen, and S. Nahavandi, “Deep reinforcement learning for multiagent systems: a review of challenges, solutions, and applications,” IEEE Transactions on Cybernetics
2020
Closest in time.
D. Isele, R. Rahimi, A. Cosgun, K. Subramanian, and K. Fujimura, “Navigating occluded intersections with autonomous vehicles using deep reinforcement learning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA)
2039
Closest in time.
G. Han, L. Xiao, and H. V. Poor, “Two-dimensional anti-jamming communication based on deep reinforcement learning,” in The 42nd IEEE International Conference on Acoustics, Speech and Signal Processing
2091
Closest in time.
H. V. Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double Q-learning,” in The Thirtieth AAAI Conference on Artificial Intelligence
2094
Closest in time.