Fetching the paper…
Reading the bibliography…
We consider the problem of multiple agents sensing and acting in environments with the goal of maximising their shared utility.
Multi-agent reinforcement learning: Independent vs. cooperative agents
M. Tan · 1993
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Introduction to reinforcement learning
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Learning communication for multi-agent systems
C. L. Giles and K. C. Jim · 2002
Earlier work this paper cites.
100 prisoners and a lightbulb
W. Wu · 2002
Earlier work this paper cites.
Cooperative multi-agent learning: The state of the art
L. Panait and S. Luke · 2005
Earlier work this paper cites.
How did language go discrete?
M. Studdert-Kennedy · 2005
Earlier work this paper cites.
Learning of communication codes in multi-agent reinforcement learning problem
T. Kasai, H. Tenmoto, and A. Kamiya · 2008
Earlier work this paper cites.
Optimal and approximate Q-value functions for decentralized POMDPs
F. A. Oliehoek, M. T. J. Spaan, and N. Vlassis · 2008
Cited alongside, same era.
Multiagent Systems: Algorithmic, Game-Theoretic, and Logical Foundations
Y. Shoham and K. Leyton-Brown · 2009
Cited alongside, same era.
QueryPOMDP: POMDP-based communication in multiagent systems
F. S. Melo, M. Spaan, and S. J. Witwicki · 2011
Cited alongside, same era.
Discovering binary codes for documents by learning deep generative models
G. Hinton and R. Salakhutdinov · 2011
Cited alongside, same era.
Coordinating multi-agent reinforcement learning with limited communication
C. Zhang and V. Lesser · 2013
Cited alongside, same era.
Empirically evaluating multiagent learning algorithms
E. Zawadzki, A. Lipson, and K. Leyton-Brown · 2014
Cited alongside, same era.
Draw: A recurrent neural network for image generation
K. Gregor, I. Danihelka, A. Graves, and D. Wierstra · 2015
Later among the works it cites.
Multiagent cooperation and competition with deep reinforcement learning
A. Tampuu, T. Matiisen, D. Kodelja, I. Kuzovkin, K. Korjus, J. Aru, J. Aru, and R. Vicente · 2015
Later among the works it cites.
Deep recurrent Q-learning for partially observable MDPs
M. Hausknecht and P. Stone · 2015
Later among the works it cites.
Language understanding for text-based games using deep reinforcement learning
K. Narasimhan, T. Kulkarni, and R. Barzilay · 2015
Later among the works it cites.
An empirical exploration of recurrent network architectures
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On the properties of neural machine translation: Encoder-decoder approaches
K. Cho, B. van Merriënboer, D. Bahdanau, and Y. Bengio · 2014
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
J. Chung, C. Gulcehre, K. Cho, and Y. Bengio · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
Cited alongside, same era.
R. Jozefowicz, W. Zaremba, and I. Sutskever · 2015
Later among the works it cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Later among the works it cites.
Multi-agent reinforcement learning as a rehearsal for decentralized planning
L. Kraemer and B. Banerjee · 2016
Closest in time.
Learning to communicate to solve riddles with deep distributed recurrent q-networks
J. N. Foerster, Y. M. Assael, N. de Freitas, and S. Whiteson · 2016
Closest in time.
BinaryNet: Training deep neural networks with weights and activations constrained to +1 or -1
M. Courbariaux and Y. Bengio · 2016
Closest in time.