TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dan Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng · 2015
Later among the works it cites.
RMSProp and equilibrated adaptive learning rates for non-convex optimization
Original
Yann N. Dauphin, Harm de Vries, Junyoung Chung, and Yoshua Bengio · 2015
Later among the works it cites.
Deep Reinforcement Learning with an Unbounded Action Space
Original
Ji He, Jianshu Chen, Xiaodong He, Jianfeng Gao, Lihong Li, Li Deng, and Mari Ostendorf · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin A. Riedmiller, Andreas Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Later among the works it cites.
Language Understanding for Text-based Games Using Deep Reinforcement Learning
Original
Karthik Narasimhan, Tejas D. Kulkarni, and Regina Barzilay · 2015
Later among the works it cites.
Understanding LSTM networks, 2015
Christopher Olah · 2015
Later among the works it cites.
Prioritized Experience Replay
Original
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2015
Later among the works it cites.
Embracing data abundance: Booktest dataset for reading comprehension
Original
Ondrej Bajgar, Rudolf Kadlec, and Jan Kleindienst · 2016
Later among the works it cites.
OpenAI Gym
Original
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Later among the works it cites.
Deep Learning
Ian Goodfellow, Yoshua Bengio, and Aaron Courville · 2016
Later among the works it cites.
Deep sentence embedding using long short-term memory networks: Analysis and application to information retrieval
Hamid Palangi, Li Deng, Yelong Shen, Jianfeng Gao, Xiaodong He, Jianshu Chen, Xinying Song, and Rabab K. Ward · 2016
Later among the works it cites.
Hybrid code networks: practical and efficient end-to-end dialog control with supervised and reinforcement learning
Original
Jason D. Williams, Kavosh Asadi, and Geoffrey Zweig · 2017
Later among the works it cites.