Fetching the paper…
Reading the bibliography…
Many tasks in AI require the collaboration of multiple agents.
Reverend bayes on inference engines: A distributed hierarchical approach
J. Pearl · 1982
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents
M. Tan · 1993
Earlier work this paper cites.
Reinforcement learning in the multi-robot domain
M. Matari · 1997
Earlier work this paper cites.
Elevator group control using multiple reinforcement learning agents
R. H. Crites and A. G. Barto · 1998
Earlier work this paper cites.
Towards collaborative and adversarial learning: A case study in robotic soccer
P. Stone and M. Veloso · 1998
Earlier work this paper cites.
Introduction to Reinforcement Learning
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Probabilistic approach to collaborative multi-robot localization
D. Fox, W. Burgard, H. Kruppa, and S. Thrun · 2000
Earlier work this paper cites.
An algorithm for distributed reinforcement learning in cooperative multi-agent systems
M. Lauer and M. A. Riedmiller · 2000
Earlier work this paper cites.
Multiagent planning with factored MDPs
C. Guestrin, D. Koller, and R. Parr · 2001
Earlier work this paper cites.
Value-function reinforcement learning in markov games
M. L. Littman · 2001
Earlier work this paper cites.
Learning communication for multi-agent systems
C. L. Giles and K. C. Jim · 2002
Earlier work this paper cites.
Reinforcement learning to play an optimal nash equilibrium in team markov games
X. Wang and T. Sandholm · 2002
Earlier work this paper cites.
Consensus and cooperation in networked multi-agent systems
R. Olfati-Saber, J. Fax, and R. Murray · 2007
Cited alongside, same era.
A comprehensive survey of multiagent reinforcement learning
L. Busoniu, R. Babuska, and B. De Schutter · 2008
Cited alongside, same era.
Learning of communication codes in multi-agent reinforcement learning problem
T. Kasai, H. Tenmoto, and A. Kamiya · 2008
Cited alongside, same era.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Cited alongside, same era.
The graph neural network model
F. Scarselli, M. Gori, A. C. Tsoi, M. Hagenbuchner, and G. Monfardini · 2009
Cited alongside, same era.
Distributed Autonomous Robotic Systems 8
P. Varshavskaya, L. P. Kaelbling, and D. Rus · 2009
Cited alongside, same era.
Gated graph sequence neural networks
Y. Li, D. Tarlow, M. Brockschmidt, and R. Zemel · 2015
Later among the works it cites.
Move evaluation in go using deep convolutional neural networks
C. J. Maddison, A. Huang, I. Sutskever, and D. Silver · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, D. Wierstra, S. Legg, and D. Hassabis · 2015
Later among the works it cites.
Towards Neural Network-based Reasoning
B. Peng, Z. Lu, H. Li, and K. Wong · 2015
Later among the works it cites.
Mazebase: A sandbox for learning from games
S. Sukhbaatar, A. Szlam, G. Synnaeve, S. Chintala, and R. Fergus · 2015
Later among the works it cites.
End-to-end memory networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Querypomdp: Pomdp-based communication in multiagent systems
F. S. Melo, M. Spaan, and S. J. Witwicki · 2011
Cited alongside, same era.
Lecture 6.5—RmsProp: Divide the gradient by a running average of its recent magnitude
T. Tieleman and G. Hinton · 2012
Cited alongside, same era.
An overview of recent progress in the study of distributed multi-agent coordination
Y. Cao, W. Yu, W. Ren, and G. Chen · 2013
Cited alongside, same era.
Coordination of communication in robot teams by reinforcement learning
D. Maravall, J. De Lope, and R. Domnguez · 2013
Cited alongside, same era.
Coordinating multi-agent reinforcement learning with limited communication
C. Zhang and V. Lesser · 2013
Cited alongside, same era.
Deep learning for real-time atari game play using offline monte-carlo tree search planning
X. Guo, S. Singh, H. Lee, R. L. Lewis, and X. Wang · 2014
Cited alongside, same era.
S. Sukhbaatar, A. Szlam, J. Weston, and R. Fergus · 2015
Later among the works it cites.
Multiagent cooperation and competition with deep reinforcement learning
A. Tampuu, T. Matiisen, D. Kodelja, I. Kuzovkin, K. Korjus, J. Aru, and R. Vicente · 2015
Later among the works it cites.
Learning to communicate to solve riddles with deep distributed recurrent Q-networks
J. N. Foerster, Y. M. Assael, N. de Freitas, and S. Whiteson · 2016
Closest in time.
Neural gpus learn algorithms
L. Kaiser and I. Sutskever · 2016
Closest in time.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Closest in time.
Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Closest in time.
Towards ai-complete question answering: A set of prerequisite toy tasks
J. Weston, A. Bordes, S. Chopra, and T. Mikolov · 2016
Closest in time.
Dynamic memory networks for visual and textual question answering
C. Xiong, S. Merity, and R. Socher · 2016
Closest in time.