Fetching the paper…
Reading the bibliography…
Recent studies have revealed that neural network-based policies can be easily fooled by adversarial examples.
X. Qu, Y.-S. Ong, Y. Hou, and X. Shen, “Memetic evolution strategy for reinforcement learning,” in 2019 IEEE Congress on Evolutionary Computation (CEC) . IEEE, 2019, pp. 1922–1928
1928
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International conference on machine learning , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
C. H. Papadimitriou and J. N. Tsitsiklis, “The complexity of markov decision processes,” Mathematics of operations research , vol. 12, no. 3, pp. 441–450, 1987
1987
Earlier work this paper cites.
J. H. Holland, “Genetic algorithms,” Scientific american , vol. 267, no. 1, pp. 66–73, 1992
1992
Earlier work this paper cites.
D. P. Bertsekas, “Nonlinear programming,” Journal of the Operational Research Society , vol. 48, no. 3, pp. 334–334, 1997
1997
Earlier work this paper cites.
Y. Jin, M. Olhofer, and B. Sendhoff, “A framework for evolutionary optimization with approximate fitness functions,” IEEE Transactions on evolutionary computation , vol. 6, no. 5, pp. 481–494, 2002
2002
Earlier work this paper cites.
Y. Jin, J. Branke et al. , “Evolutionary optimization in uncertain environments-a survey,” IEEE Transactions on evolutionary computation , vol. 9, no. 3, pp. 303–317, 2005
2005
Earlier work this paper cites.
Y. Jin, “A comprehensive survey of fitness approximation in evolutionary computation,” Soft computing , vol. 9, no. 1, pp. 3–12, 2005
2005
Earlier work this paper cites.
B. Liu, L. Wang, Y.-H. Jin, F. Tang, and D.-X. Huang, “Improved particle swarm optimization combined with chaos,” Chaos, Solitons & Fractals , vol. 25, no. 5, pp. 1261–1271, 2005
2005
Earlier work this paper cites.
B. Liu, L. Wang, and Y.-H. Jin, “An effective pso-based memetic algorithm for flow shop scheduling,” IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics) , vol. 37, no. 1, pp. 18–27, 2007
2007
Earlier work this paper cites.
K. Deep, K. P. Singh, M. L. Kansal, and C. Mohan, “A real coded genetic algorithm for solving integer and mixed integer optimization problems,” Applied Mathematics and Computation , vol. 212, no. 2, pp. 505–518, 2009
2009
Earlier work this paper cites.
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International Conference on Machine Learning , 2015, pp. 1889–1897
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, p. 529, 2015
2015
Earlier work this paper cites.
2016
Cited alongside, same era.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of go with deep neural networks and tree search,” nature , vol. 529, no. 7587, p. 484, 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
Z. Li, J. Liu, Z. Huang, Y. Peng, H. Pu, and L. Ding, “Adaptive impedance control of human–robot cooperation using reinforcement learning,” IEEE Transactions on Industrial Electronics , vol. 64, no. 10, pp. 8013–8022, 2017
2017
Cited alongside, same era.
I. Goodfellow, N. Papernot, S. Huang, Y. Duan, P. Abbeel, and J. Clark, “Attacking machine learning with adversarial examples,” OpenAI. https://blog. openai. com/adversarial-example-research , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
X. Ma, Q. Zhang, G. Tian, J. Yang, and Z. Zhu, “On tchebycheff decomposition approaches for multiobjective evolutionary optimization,” IEEE Transactions on Evolutionary Computation , vol. 22, no. 2, pp. 226–244, 2017
2017
Later among the works it cites.
L. Feng, Y.-S. Ong, S. Jiang, and A. Gupta, “Autoencoding evolutionary search with learning across heterogeneous problems,” IEEE Transactions on Evolutionary Computation , vol. 21, no. 5, pp. 760–772, 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2017
Cited alongside, same era.
P.-Y. Chen, H. Zhang, Y. Sharma, J. Yi, and C.-J. Hsieh, “Zoo: Zeroth order optimization based black-box attacks to deep neural networks without training substitute models,” in Proceedings of the 10th ACM Workshop on Artificial Intelligence and Security . ACM, 2017, pp. 15–26
2017
Cited alongside, same era.
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta, “Robust adversarial reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning-Volume 70 . JMLR. org, 2017, pp. 2817–2826
2017
Cited alongside, same era.
2017
Cited alongside, same era.
S. Parbhoo, J. Bogojeska, M. Zazzi, V. Roth, and F. Doshi-Velez, “Combining kernel and model based learning for hiv therapy selection,” AMIA Summits on Translational Science Proceedings , vol. 2017, p. 239, 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Y. Wu, E. Mansimov, R. B. Grosse, S. Liao, and J. Ba, “Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation,” in Advances in neural information processing systems , 2017, pp. 5279–5288
2017
Cited alongside, same era.
A. RAFFIN, “Reinforcement learning baseline zoo,” https://https://github.com/araffin/rl-baselines-zoo
2017
Later among the works it cites.
Z. Xie and Y. Jin, “An extended reinforcement learning framework to model cognitive development with enactive pattern representation,” IEEE Transactions on Cognitive and Developmental Systems , vol. 10, no. 3, pp. 738–750, 2018
2018
Later among the works it cites.
F. de La Bourdonnaye, C. Teulière, J. Triesch, and T. Chateau, “Learning to touch objects through stage-wise deep reinforcement learning,” in 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2018, pp. 1–9
2018
Later among the works it cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Later among the works it cites.
Y. Jin, H. Wang, T. Chugh, D. Guo, and K. Miettinen, “Data-driven evolutionary optimization: an overview and case studies,” IEEE Transactions on Evolutionary Computation , vol. 23, no. 3, pp. 442–458, 2018
2018
Later among the works it cites.
X. Yuan, P. He, Q. Zhu, and X. Li, “Adversarial examples: Attacks and defenses for deep learning,” IEEE transactions on neural networks and learning systems , 2019
2019
Closest in time.
A. Raghu, “Reinforcement learning for sepsis treatment: Baselines and analysis,” Reinforcement Learning for Real Life , 2019
2019
Closest in time.
2019
Closest in time.
Z. Zhang, Y.-S. Ong, D. Wang, and B. Xue, “A collaborative multiagent reinforcement learning method based on policy gradient potential,” IEEE transactions on cybernetics , 2019
2019
Closest in time.
J. Su, D. V. Vargas, and K. Sakurai, “One pixel attack for fooling deep neural networks,” IEEE Transactions on Evolutionary Computation , 2019
2019
Closest in time.
A. Pattanaik, Z. Tang, S. Liu, G. Bommannan, and G. Chowdhary, “Robust deep reinforcement learning with adversarial attacks,” in Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems . International Foundation for Autonomous Agents and Multiagent Systems, 2018, pp. 2040–2042
2042
Closest in time.