Fetching the paper…
Reading the bibliography…
Several approaches have recently been proposed for learning decentralized deep multiagent policies that coordinate via a differentiable communication channel.
Denotational semantics for agent communication language
Frank Guerin and Jeremy Pitt. 2001 · 2001
Earlier work this paper cites.
The complexity of decentralized control of Markov decision processes
Daniel S Bernstein, Robert Givan, Neil Immerman, and Shlomo Zilberstein. 2002 · 2002
Earlier work this paper cites.
Reasoning about joint beliefs for execution-time communication decisions
Maayan Roth, Reid Simmons, and Manuela Veloso. 2005 · 2005
Earlier work this paper cites.
Informative communication in word production and word learning
Michael C Frank, Noah D Goodman, Peter Lai, and Joshua B Tenenbaum. 2009 · 2009
Earlier work this paper cites.
Caltech-UCSD Birds 200
P. Welinder, S. Branson, T. Mita, C. Wah, F. Schroff, S. Belongie, and P. Perona. 2010 · 2010
Earlier work this paper cites.
Generating legible motion
Anca Dragan and Siddhartha Srinivasa. 2013 · 2013
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
Kyunghyun Cho, Bart van Merriënboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
ReferItGame: Referring to objects in photographs of natural scenes
Sahar Kazemzadeh, Vicente Ordonez, Mark Matten, and Tamara L Berg. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc VV Le. 2014 · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Matthew D Zeiler and Rob Fergus. 2014 · 2014
Cited alongside, same era.
Deep recurrent q-learning for partially observable mdps
Matthew Hausknecht and Peter Stone. 2015 · 2015
Cited alongside, same era.
A Bayesian model of grounded color semantics
Brian McMahan and Matthew Stone. 2015 · 2015
Cited alongside, same era.
Reasoning about pragmatics with neural listeners and speakers
Jacob Andreas and Dan Klein. 2016 · 2016
Cited alongside, same era.
Optimally solving Dec-POMDPs as continuous-state MDPs
Jilles Steeve Dibangoye, Christopher Amato, Olivier Buffet, and François Charpillet. 2016 · 2016
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning
Jakob Foerster, Yannis M Assael, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Inference networks for sequential monte carlo in graphical models
Brooks Paige and Frank Wood. 2016 · 2016
Later among the works it cites.
Inferring logical forms from denotations
Panupong Pasupat and Percy Liang. 2016 · 2016
Later among the works it cites.
Learning deep representations of fine-grained visual descriptions
Scott Reed, Zeynep Akata, Honglak Lee, and Bernt Schiele. 2016 · 2016
Later among the works it cites.
Why should I trust you?: Explaining the predictions of any classifier
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Later among the works it cites.
Visual analysis of hidden state dynamics in recurrent neural networks
Hendrik Strobelt, Sebastian Gehrmann, Bernd Huber, Hanspeter Pfister, and Alexander M Rush. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Compact bilinear pooling
Yang Gao, Oscar Beijbom, Ning Zhang, and Trevor Darrell. 2016 · 2016
Cited alongside, same era.
A paradigm for situated and goal-driven language learning
Jon Gauthier and Igor Mordatch. 2016 · 2016
Cited alongside, same era.
Generating visual explanations
Lisa Anne Hendricks, Zeynep Akata, Marcus Rohrbach, Jeff Donahue, Bernt Schiele, and Trevor Darrell. 2016 · 2016
Cited alongside, same era.
Learning to generate compositional color descriptions
Will Monroe, Noah D Goodman, and Christopher Potts. 2016 · 2016
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni. 2016a
Cited in the paper.
Towards multi-agent communication-based language learning
Angeliki Lazaridou, Nghia The Pham, and Marco Baroni. 2016b
Cited in the paper.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, Rob Fergus, et al. 2016 · 2016
Later among the works it cites.
Proceedings of nips 2016 workshop on interpretable machine learning for complex systems
Andrew Gordon Wilson, Been Kim, and William Herlands. 2016 · 2016
Later among the works it cites.
A joint speaker-listener-reinforcer model for referring expressions
Licheng Yu, Hao Tan, Mohit Bansal, and Tamara L Berg. 2016 · 2016
Later among the works it cites.
Context-aware captions from context-agnostic supervision
Ramakrishna Vedantam, Samy Bengio, Kevin Murphy, Devi Parikh, and Gal Chechik. 2017 · 2017
Closest in time.