Fetching the paper…
Reading the bibliography…
One of the long-term goals of artificial intelligence is to build an agent that can communicate intelligently with human in natural language.
Verbal Behavior
B. F. Skinner · 1957
Earlier work this paper cites.
Reinforcement Learning: An Introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Knowledge and Learning in Natural Language
C. D. Yang · 2003
Earlier work this paper cites.
Early language acquisition: cracking the speech code
P. K. Kuhl · 2004
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
J. C. Duchi, E. Hazan, and Y. Singer · 2011
Earlier work this paper cites.
Playing Atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
Sequence to sequence learning with neural networks
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Deep captioning with multimodal recurrent neural networks (m-RNN)
J. Mao, W. Xu, Y. Yang, J. Wang, Z. Huang, and A. Yuille · 2015
Earlier work this paper cites.
O. Vinyals and Q. V. Le · 2015
Earlier work this paper cites.
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2015
Cited alongside, same era.
Semantically conditioned LSTM-based natural language generation for spoken dialogue systems
T. Wen, M. Gasic, N. Mrksic, P. Su, D. Vandyke, and S. J. Young · 2015
Cited alongside, same era.
Learning to learn by gradient descent by gradient descent
M. Andrychowicz, M. Denil, S. G. Colmenarejo, M. W. Hoffman, D. Pfau, T. Schaul, and N. de Freitas · 2016
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning
J. N. Foerster, Y. M. Assael, N. de Freitas, and S. Whiteson · 2016
Cited alongside, same era.
Deep reinforcement learning with a natural language action space
J. He, J. Chen, X. He, J. Gao, L. Li, L. Deng, and M. Ostendorf · 2016
Cited alongside, same era.
Deep reinforcement learning for dialogue generation
J. Li, W. Monroe, A. Ritter, D. Jurafsky, M. Galley, and J. Gao · 2016
Dialog-based language learning
J. Weston · 2016
Later among the works it cites.
Video paragraph captioning using hierarchical recurrent neural networks
H. Yu, J. Wang, Z. Huang, Y. Yang, and W. Xu · 2016
Later among the works it cites.
An actor-critic algorithm for sequence prediction
D. Bahdanau, P. Brakel, K. Xu, A. Goyal, R. Lowe, J. Pineau, A. C. Courville, and Y. Bengio · 2017
Closest in time.
Learning cooperative visual dialog agents with deep reinforcement learning
A. Das, S. Kottur, , J. M. Moura, S. Lee, and D. Batra · 2017
Closest in time.
Multi-agent cooperation and the emergence of (natural) language
A. Lazaridou, A. Peysakhovich, and M. Baroni · 2017
Closest in time.
Emergence of grounded compositional language in multi-agent populations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Reinforcement contingencies in language acquisition
A. I. Petursdottir and J. R. Mellor · 2016
Cited alongside, same era.
Sequence level training with recurrent neural networks
M. Ranzato, S. Chopra, M. Auli, and W. Zaremba · 2016
Cited alongside, same era.
Building end-to-end dialogue systems using generative hierarchical neural network models
I. V. Serban, A. Sordoni, Y. Bengio, A. C. Courville, and J. Pineau · 2016
Cited alongside, same era.
Learning multiagent communication with backpropagation
S. Sukhbaatar, A. Szlam, and R. Fergus · 2016
Cited alongside, same era.
Learning through dialogue interactions
J. Li, A. H. Miller, S. Chopra, M. Ranzato, and J. Weston
Cited in the paper.
Adversarial learning for neural dialogue generation
J. Li, W. Monroe, T. Shi, A. Ritter, and D. Jurafsky
Cited in the paper.
I. Mordatch and P. Abbeel · 2017
Closest in time.
Third-person imitation learning
B. C. Stadie, P. Abbeel, and I. Sutskever · 2017
Closest in time.
End-to-end optimization of goal-driven and visually grounded dialogue systems
F. Strub, H. de Vries, J. Mary, B. Piot, A. C. Courville, and O. Pietquin · 2017
Closest in time.
SeqGAN: Sequence generative adversarial nets with policy gradient
L. Yu, W. Zhang, J. Wang, and Y. Yu · 2017
Closest in time.