Fetching the paper…
Reading the bibliography…
We present a method for combining multi-agent communication and traditional data-driven approaches to natural language learning, with an end goal of teaching agents to communicate with humans in natural language.
Emergent linguistic phenomena in multi-agent communication games
Laura Graesser, Kyunghyun Cho, and Douwe Kiela. 2019 · 1901
Earlier work this paper cites.
Anti-efficient encoding in emergent communication
Rahma Chaabouni, Eugene Kharitonov, Emmanuel Dupoux, and Marco Baroni. 2019 · 1905
Earlier work this paper cites.
Fine-tuning language models from human preferences
Daniel M Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving. 2019 · 1909
Earlier work this paper cites.
Philosophical investigations
Ludwig Wittgenstein. 1953 · 1953
Earlier work this paper cites.
How to do things with words
John Langshaw Austin. 1975 · 1975
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Improved backing-off for m-gram language modeling
R. Kneser and H. Ney. 1995 · 1995
Earlier work this paper cites.
Conceptual pacts and lexical choice in conversation
Susan E Brennan and Herbert H Clark. 1996 · 1996
Earlier work this paper cites.
Using language
Herbert H Clark. 1996 · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Building applied natural language generation systems
Ehud Reiter and Robert Dale. 1997 · 1997
Earlier work this paper cites.
On the interaction between supervision and self-play in emergent communication
Ryan Lowe, Abhinav Gupta, Jakob N. Foerster, Douwe Kiela, and Joelle Pineau. 2020 · 2002
Earlier work this paper cites.
Countering language drift with seeded iterated learning
Yuchen Lu, Soumye Singhal, Florian Strub, Olivier Pietquin, and Aaron C. Courville. 2020 · 2003
Earlier work this paper cites.
A survey of statistical user simulation techniques for reinforcement-learning of dialogue management strategies
Jost Schatzmann, Karl Weilhammer, Matt Stuttle, and Steve Young. 2006 · 2006
Cited alongside, same era.
Recurrent neural network based language model
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur. 2010 · 2010
Cited alongside, same era.
Bringing semantics into focus using visual abstraction
C Lawrence Zitnick and Devi Parikh. 2013 · 2013
Cited alongside, same era.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Cited alongside, same era.
Emergent communication in a multi-modal, multi-step referential game
Katrina Evtimova, Andrew Drozdov, Douwe Kiela, and Kyunghyun Cho. 2017 · 2017
Later among the works it cites.
Emergence of language with multi-agent games: Learning to communicate with sequences of symbols
Serhii Havrylov and Ivan Titov. 2017 · 2017
Later among the works it cites.
Natural language does not emerge’naturally’in multi-agent dialog
Satwik Kottur, José MF Moura, Stefan Lee, and Dhruv Batra. 2017 · 2017
Later among the works it cites.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni. 2017 · 2017
Later among the works it cites.
Context-aware captions from context-agnostic supervision
Ramakrishna Vedantam, Samy Bengio, Kevin Murphy, Devi Parikh, and Gal Chechik. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Will Monroe and Christopher Potts. 2015 · 2015
Cited alongside, same era.
Oriol Vinyals and Quoc Le. 2015 · 2015
Cited alongside, same era.
Reasoning about pragmatics with neural listeners and speakers
Jacob Andreas and Dan Klein. 2016 · 2016
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning
Jakob Foerster, Ioannis Alexandros Assael, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Seeing through the human reporting bias: Visual classifiers from noisy human-centric labels
Ishan Misra, C. Zitnick, Margaret Mitchell, and Ross Girshick. 2016 · 2016
Cited alongside, same era.
Yfcc100m: the new data in multimedia research
Bart Thomee, David A. Shamma, Gerald Friedland, Benjamin Elizalde, Karl Ni, Douglas Poland, Damian Borth, and Li-Jia Li. 2016 · 2016
Cited alongside, same era.
How agents see things: On visual representations in an emergent language game
Diane Bouchacourt and Marco Baroni. 2018 · 2018
Later among the works it cites.
Pragmatically informative image captioning with character-level inference
Reuben Cohn-Gordon, Noah Goodman, and Christopher Potts. 2018 · 2018
Later among the works it cites.
Speaker-follower models for vision-and-language navigation
Daniel Fried, Ronghang Hu, Volkan Cirik, Anna Rohrbach, Jacob Andreas, Louis-Philippe Morency, Taylor Berg-Kirkpatrick, Kate Saenko, Dan Klein, and Trevor Darrell. 2018 · 2018
Later among the works it cites.
Emergence of linguistic communication from referential games with symbolic and pixel input
Angeliki Lazaridou, Karl Moritz Hermann, Karl Tuyls, and Stephen Clark. 2018 · 2018
Later among the works it cites.
Bootstrapping a neural conversational agent with dialogue self-play, crowdsourcing and on-line reinforcement learning
Pararth Shah, Dilek Hakkani-Tür, Bing Liu, and Gokhan Tür. 2008 · 2018
Later among the works it cites.
On the utility of learning about humans for human-ai coordination
Micah Carroll, Rohin Shah, Mark K Ho, Tom Griffiths, Sanjit Seshia, Pieter Abbeel, and Anca Dragan. 2019 · 2019
Later among the works it cites.
Countering language drift via visual grounding
Jason Lee, Kyunghyun Cho, and Douwe Kiela. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.