Human-Level Control through Deep Reinforcement Learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin A. Riedmiller, Andreas Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis. 2015 · 2015
Cited alongside, same era.
Overcoming Catastrophic Forgetting in Neural Networks
Original
James Kirkpatrick, Razvan Pascanu, Neil C. Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A. Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell. 2016 · 2016
Cited alongside, same era.
How Transferable are Neural Networks in NLP Applications?. In Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP) . 479–489
Lili Mou, Zhao Meng, Rui Yan, Ge Li, Yan Xu, Lu Zhang, and Zhi Jin. 2016 · 2016
Cited alongside, same era.
Actor-Mimic: Deep Multitask and Transfer Reinforcement Learning. In Proceedings of the International Conference on Learning Representations (ICLR)
Emilio Parisotto, Lei Jimmy Ba, and Ruslan Salakhutdinov. 2016 · 2016
Cited alongside, same era.
Progressive Neural Networks
Original
Andrei A. Rusu, Neil C. Rabinowitz, Guillaume Desjardins, Hubert Soyer, James Kirkpatrick, Koray Kavukcuoglu, Razvan Pascanu, and Raia Hadsell. 2016 · 2016
Cited alongside, same era.
Mastering the Game of Go with Deep Neural Networks and Tree Search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Vedavyas Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy P. Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis. 2016 · 2016
Cited alongside, same era.
Hindsight Experience Replay. In Advances in Neural Information Processing Systems (NeurIPS) . 5048–5058
Marcin Andrychowicz, Dwight Crow, Alex Ray, Jonas Schneider, Rachel Fong, Peter Welinder, Bob McGrew, Josh Tobin, Pieter Abbeel, and Wojciech Zaremba. 2017 · 2017
Cited alongside, same era.
Sharp Minima Can Generalize For Deep Nets. In Proceedings of the International Conference on Machine Learning (ICML) . 1019–1028
Laurent Dinh, Razvan Pascanu, Samy Bengio, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks. In Proceedings of the International Conference on Machine Learning (ICML) . 1126–1135
Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017 · 2017
Cited alongside, same era.
Towards Generalization and Simplicity in Continuous Control. In Advances in Neural Information Processing Systems (NeurIPS) . 6550–6561
Aravind Rajeswaran, Kendall Lowrey, Emanuel Todorov, and Sham M. Kakade. 2017 · 2017
Cited alongside, same era.
Distral: Robust Multitask Reinforcement Learning. In Advances in Neural Information Processing Systems (NeurIPS) . 4496–4506
Yee Whye Teh, Victor Bapst, Wojciech M. Czarnecki, John Quan, James Kirkpatrick, Raia Hadsell, Nicolas Heess, and Razvan Pascanu. 2017 · 2017
Cited alongside, same era.
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures. In Proceedings of the International Conference on Machine Learning (ICML) . 1406–1415
Lasse Espeholt, Hubert Soyer, Rémi Munos, Karen Simonyan, Volodymyr Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, Shane Legg, and Koray Kavukcuoglu. 2018 · 2018
Cited alongside, same era.