Fetching the paper…
Reading the bibliography…
State-of-the-art meta reinforcement learning algorithms typically assume the setting of a single agent interacting with its environment in a sequential manner.
Asynchronous methods for deep reinforcement learning. In International conference on machine learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Yoshua Bengio, Patrice Simard, and Paolo Frasconi. 1994 · 1994
Earlier work this paper cites.
A Bayesian Framework for Reinforcement Learning. In Proceedings of the Seventeenth International Conference on Machine Learning
Malcolm JA Strens. 2000 · 2000
Earlier work this paper cites.
Concurrent Reinforcement Learning from Customer Interactions. In Proceedings of the 30th International Conference on Machine Learning
David Silver, Leonard Newnham, David Barker, Suzanne Weller, and Jason McFall. 2013 · 2013
Earlier work this paper cites.
Sequence to sequence learning with neural networks. In Advances in neural information processing systems
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Earlier work this paper cites.
Charles Blundell, Benigno Uria, Alexander Pritzel, Yazhe Li, Avraham Ruderman, Joel Z Leibo, Jack Rae, Daan Wierstra, and Demis Hassabis. 2016 · 2016
Earlier work this paper cites.
RL2: Fast Reinforcement Learning via Slow Reinforcement Learning
Yan Duan, John Schulman, Xi Chen, Peter L Bartlett, Ilya Sutskever, and Pieter Abbeel. 2016 · 2016
Earlier work this paper cites.
Learning to communicate with deep multi-agent reinforcement learning. In Advances in Neural Information Processing Systems
Jakob Foerster, Ioannis Alexandros Assael, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Earlier work this paper cites.
Learning multiagent communication with backpropagation. In Advances in Neural Information Processing Systems
Sainbayar Sukhbaatar, Rob Fergus, et al · 2016
Cited alongside, same era.
Learning to reinforcement learn
Jane X Wang, Zeb Kurth-Nelson, Dhruva Tirumala, Hubert Soyer, Joel Z Leibo, Remi Munos, Charles Blundell, Dharshan Kumaran, and Matt Botvinick. 2016 · 2016
Cited alongside, same era.
Continuous adaptation via meta-learning in nonstationary and competitive environments
Maruan Al-Shedivat, Trapit Bansal, Yuri Burda, Ilya Sutskever, Igor Mordatch, and Pieter Abbeel. 2017 · 2017
Cited alongside, same era.
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks. In International Conference on Machine Learning
C. Finn, P. Abbeel, and S. Levine. 2017 · 2017
Cited alongside, same era.
Vain: Attentional multi-agent predictive modeling. In Advances in Neural Information Processing Systems
Scalable Coordinated Exploration in Concurrent Reinforcement Learning. In Advances in Neural Information Processing Systems
Maria Dimakopoulou, Ian Osband, and Benjamin Van Roy. 2018 · 2018
Later among the works it cites.
Coordinated Exploration in Concurrent Reinforcement Learning. In International Conference on Machine Learning
Maria Dimakopoulou and Benjamin Van Roy. 2018 · 2018
Later among the works it cites.
Unsupervised Meta-Learning for Reinforcement Learning
Abhishek Gupta, Benjamin Eysenbach, Chelsea Finn, and Sergey Levine. 2018a · 2018
Later among the works it cites.
Meta-Reinforcement Learning of Structured Exploration Strategies
Abhishek Gupta, Russell Mendonca, YuXuan Liu, Pieter Abbeel, and Sergey Levine. 2018b · 2018
Later among the works it cites.
Modeling Others using Oneself in Multi-Agent Reinforcement Learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yedid Hoshen. 2017 · 2017
Cited alongside, same era.
Deepstack: Expert-level artificial intelligence in heads-up no-limit poker
M. Moravčík, M. Schmid, N. Burch, V. Lisỳ, D. Morrill, N. Bard, T. Davis, K. Waugh, M. Johanson, and M. Bowling. 2017 · 2017
Cited alongside, same era.
Neural Episodic Control. In Proceedings of the 34th International Conference on Machine Learning
Alexander Pritzel, Benigno Uria, Sriram Srinivasan, Adrià Puigdomènech Badia, Oriol Vinyals, Demis Hassabis, Daan Wierstra, and Charles Blundell. 2017 · 2017
Cited alongside, same era.
Mastering the game of Go without human knowledge
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al · 2017
Cited alongside, same era.
Roberta Raileanu, Emily Denton, Arthur Szlam, and Rob Fergus. 2018 · 2018
Later among the works it cites.
Been There, Done That: Meta-Learning with Episodic Recall
Samuel Ritter, Jane X Wang, Zeb Kurth-Nelson, Siddhant M Jayakumar, Charles Blundell, Razvan Pascanu, and Matthew Botvinick. 2018 · 2018
Later among the works it cites.
Some considerations on learning to explore via meta-reinforcement learning
Bradly C Stadie, Ge Yang, Rein Houthooft, Xi Chen, Yan Duan, Yuhuai Wu, Pieter Abbeel, and Ilya Sutskever. 2018 · 2018
Later among the works it cites.