Fetching the paper…
Reading the bibliography…
$\textit{No man is an island.}$ Humans communicate with a large community by coordinating with different interlocutors within short conversations.
ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks
Shridhar, M., Thomason, J., Gordon, D., Bisk, Y., Han, W., Mottaghi, R., Zettlemoyer, L., and Fox, D · 1912
Earlier work this paper cites.
Philosophical Investigations
Wittgenstein, L · 1953
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
Premack, D. and Woodruff, G · 1978
Earlier work this paper cites.
Conceptual pacts and lexical choice in conversation
Brennan, S. E. and Clark, H. H · 1996
Earlier work this paper cites.
Reinforcement learning: A survey
Kaelbling, L. P., Littman, M. L., and Moore, A. W · 1996
Earlier work this paper cites.
Computational simulations of the emergence of grammar
Batali, J · 1998
Earlier work this paper cites.
Neural systems involved in’theory of mind’
Siegal, M. and Varley, R · 2002
Earlier work this paper cites.
Progress in the simulation of emergent communication and language
Wagner, K., Reggia, J. A., Uriagereka, J., and Wilkinson, G. S · 2003
Earlier work this paper cites.
Age-of-acquisition effects in word and picture identification
Juhasz, B. J · 2005
Earlier work this paper cites.
The interface of language and theory of mind
De Villiers, J · 2007
Earlier work this paper cites.
ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
Shridhar, M., Yuan, X., Côté, M.-A., Bisk, Y., Trischler, A., and Hausknecht, M · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G., and Bagnell, D · 2011
Earlier work this paper cites.
Predicting pragmatic reasoning in language games
Frank, M. C. and Goodman, N. D · 2012
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., and Zitnick, C. L · 2014
Earlier work this paper cites.
Learning in the rational speech acts model
Monroe, W. and Potts, C · 2015
Earlier work this paper cites.
The computational theory of mind
Rescorla, M · 2015
Earlier work this paper cites.
Learning and policy search in stochastic dynamical systems with bayesian neural networks
Depeweg, S., Hernández-Lobato, J. M., Doshi-Velez, F., and Udluft, S · 2016
Cited alongside, same era.
Improving pilco with bayesian neural network dynamics models
Gal, Y., McAllister, R., and Rasmussen, C. E · 2016
Cited alongside, same era.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Lazaridou, A., Peysakhovich, A., and Baroni, M · 2016
Cited alongside, same era.
Learning language games through interaction
Wang, S. I., Liang, P., and Manning, C · 2016
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S · 2017
How to train your MAML
Antoniou, A., Edwards, H., and Storkey, A · 2019
Later among the works it cites.
How efficiency shapes human language
Gibson, E., Futrell, R., Piantadosi, S. P., Dautriche, I., Mahowald, K., Bergen, L., and Levy, R · 2019
Later among the works it cites.
When to trust your model: Model-based policy optimization
Janner, M., Fu, J., Zhang, M., and Levine, S · 2019
Later among the works it cites.
Revisiting the evaluation of theory of mind through question answering
Le, M., Boureau, Y.-L., and Nickel, M · 2019
Later among the works it cites.
Ease-of-teaching and language structure from emergent communication
Li, F. and Bowling, M · 2019
Later among the works it cites.
Pressure to communicate across knowledge asymmetries leads to pedagogically supportive language input
Morris, B. and Yurovsky, D · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Categorical Reparameterization with Gumbel-Softmax
Jang, E., Gu, S., and Poole, B · 2017
Cited alongside, same era.
How agents see things: On visual representations in an emergent language game
Bouchacourt, D. and Baroni, M · 2018
Cited alongside, same era.
Emergent communication through negotiation
Cao, K., Lazaridou, A., Lanctot, M., Leibo, J. Z., Tuyls, K., and Clark, S · 2018
Cited alongside, same era.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
Chua, K., Calandra, R., McAllister, R., and Levine, S · 2018
Cited alongside, same era.
Unified pragmatic models for generating and following instructions
Fried, D., Andreas, J., and Klein, D · 2018
Cited alongside, same era.
Acoustic-prosodic and lexical entrainment in deceptive dialogue
Levitan, S. I., Xiang, J., and Hirschberg, J · 2018
Cited alongside, same era.
Shaping representations through communication: community size effect in artificial learning systems
Tieleman, O., Lazaridou, A., Mourad, S., Blundell, C., and Precup, D · 2019
Later among the works it cites.
Exploring zero-shot emergent communication in embodied multi-agent populations
Bullard, K., Meier, F., Kiela, D., Pineau, J., and Foerster, J · 2020
Later among the works it cites.
Compositionality and generalization in emergent languages
Chaabouni, R., Kharitonov, E., Bouchacourt, D., Dupoux, E., and Baroni, M · 2020
Later among the works it cites.
Compositionality and capacity in emergent languages
Gupta, A., Resnick, C., Foerster, J., Dai, A., and Cho, K · 2020
Later among the works it cites.
“Other-Play” for Zero-Shot Coordination
Hu, H., Lerer, A., Peysakhovich, A., and Foerster, J · 2020
Later among the works it cites.
Entropy minimization in emergent languages
Kharitonov, E., Chaabouni, R., Bouchacourt, D., and Baroni, M · 2020
Later among the works it cites.
A mathematical theory of cooperative communication
Wang, P., Wang, J., Paranamana, P., and Shafto, P · 2020
Later among the works it cites.
Emergence of pragmatics from referential game between theory of mind agents
Yuan, L., Fu, Z., Shen, J., Xu, L., Shen, J., and Zhu, S.-C · 2020
Later among the works it cites.
Neural recursive belief states in multi-agent reinforcement learning
Moreno, P., Hughes, E., McKee, K. R., Pires, B. A., and Weber, T · 2021
Closest in time.