Fetching the paper…
Reading the bibliography…
Pretraining on human corpus and then finetuning in a simulator has become a standard pipeline for training a goal-oriented dialogue agent.
Convention: A Philosophical Study
Lewis, D. K · 1969
Earlier work this paper cites.
The second naive physics manifesto
Hayes, P. J · 1988
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
McCloskey, M. and Cohen, N. J · 1989
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J · 1992
Earlier work this paper cites.
Building a large annotated corpus of english: The penn treebank
Marcus, M., Santorini, B., and Marcinkiewicz, M. A · 1993
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
A stochastic model of human-machine interaction for learning dialog strategies
Levin, E., Pieraccini, R., and Eckert, W · 2000
Earlier work this paper cites.
Spontaneous evolution of linguistic structure-an iterated learning model of the emergence of regularity and irregularity
Kirby, S · 2001
Earlier work this paper cites.
Natural language from artificial life
Kirby, S · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J · 2002
Earlier work this paper cites.
The task rehearsal method of life-long learning: Overcoming impoverished data
Silver, D. L. and Mercer, R. E · 2002
Earlier work this paper cites.
A neural probabilistic language model
Bengio, Y., Ducharme, R., Vincent, P., and Jauvin, C · 2003
Earlier work this paper cites.
A bayesian view of language evolution by iterated learning
Griffiths, T. L. and Kalish, M. L · 2005
Earlier work this paper cites.
A survey of statistical user simulation techniques for reinforcement-learning of dialogue management strategies
Schatzmann, J., Weilhammer, K., Stuttle, M., and Young, S · 2006
Earlier work this paper cites.
Iterated learning: Intergenerational knowledge transmission reveals inductive biases
Kalish, M. L., Griffiths, T. L., and Lewandowsky, S · 2007
Earlier work this paper cites.
Moses: Open source toolkit for statistical machine translation
Koehn, P., Hoang, H., Birch, A., Callison-Burch, C., Federico, M., Bertoldi, N., Cowan, B., Shen, W., Moran, C., Zens, R., et al · 2007
Earlier work this paper cites.
Towards incremental speech generation in dialogue systems
Skantze, G. and Hjalmarsson, A · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Collobert, R., Weston, J., Bottou, L., Karlen, M., Kavukcuoglu, K., and Kuksa, P · 2011
Earlier work this paper cites.
Wit3: Web inventory of transcribed and translated talks
Cettolo, M., Girardi, C., and Federico, M · 2012
Earlier work this paper cites.
Data-driven methods for adaptive spoken dialogue systems: Computational learning for conversational interfaces
Lemon, O. and Pietquin, O · 2012
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Cited alongside, same era.
Iterated learning and the evolution of language
Kirby, S., Griffiths, T., and Smith, K · 2014
Cited alongside, same era.
Microsoft coco: Common objects in context
Lin, T.-Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., and Zitnick, C. L · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y · 2015
Cited alongside, same era.
Multi30k: Multilingual english-german image descriptions
Elliott, D., Frank, S., Sima’an, K., and Specia, L · 2016
Cited alongside, same era.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Interactive reinforcement learning for object grounding via self-talking
Zhu, Y., Zhang, S., and Metaxas, D · 2017
Later among the works it cites.
Speaker-follower models for vision-and-language navigation
Fried, D., Hu, R., Cirik, V., Rohrbach, A., Andreas, J., Morency, L.-P., Berg-Kirkpatrick, T., Saenko, K., Klein, D., and Darrell, T · 2018
Later among the works it cites.
Unsupervised machine translation using monolingual corpora only
Lample, G., Conneau, A., Denoyer, L., and Ranzato, M · 2018
Later among the works it cites.
Airdialogue: An environment for goal-oriented dialogue research
Wei, W., Le, Q., Dai, A., and Li, J · 2018
Later among the works it cites.
Community regularization of visually-grounded dialog
Agarwal, A., Gurumurthy, S., Sharma, V., Lewis, M., and Sycara, K · 2019
Later among the works it cites.
Systematic generalization: What is required and can it be learned?
Bahdanau, D., Murty, S., Noukhovitch, M., Nguyen, T. H., de Vries, H., and Courville, A · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Lazaridou, A., Peysakhovich, A., and Baroni, M · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Sennrich, R., Haddow, B., and Birch, A · 2016
Cited alongside, same era.
Learning end-to-end goal-oriented dialog
Bordes, A., Boureau, Y.-L., and Weston, J · 2017
Cited alongside, same era.
Evaluating visual conversational agents via cooperative human-ai games
Chattopadhyay, P., Yadav, D., Prabhu, V., Chandrasekaran, A., Das, A., Lee, S., Batra, D., and Parikh, D · 2017
Cited alongside, same era.
Self-sustaining iterated learning
Chazelle, B. and Wang, C · 2017
Cited alongside, same era.
Learning cooperative visual dialog agents with deep reinforcement learning
Das, A., Kottur, S., Moura, J. M., Lee, S., and Batra, D · 2017
Cited alongside, same era.
Later among the works it cites.
Iterated learning in dynamic social networks
Chazelle, B. and Wang, C · 2019
Later among the works it cites.
Emergence of compositional language with deep generational transmission
Cogswell, M., Lu, J., Lee, S., Parikh, D., and Batra, D · 2019
Later among the works it cites.
Neural approaches to conversational ai
Gao, J., Galley, M., Li, L., et al · 2019
Later among the works it cites.
Guo, S., Ren, Y., Havrylov, S., Frank, S., Titov, I., and Smith, K · 2019
Later among the works it cites.
Seeded self-play for language learning
Gupta, A., Lowe, R., Foerster, J., Kiela, D., and Pineau, J · 2019
Later among the works it cites.
Countering language drift via visual grounding
Lee, J., Cho, K., and Kiela, D · 2019
Later among the works it cites.
Ease-of-teaching and language structure from emergent communication
Li, F. and Bowling, M · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., and Sutskever, I · 2019
Later among the works it cites.
Self-training with noisy student improves imagenet classification
Xie, Q., Hovy, E., Luong, M.-T., and Le, Q. V · 2019
Later among the works it cites.
Towards a human-like open-domain chatbot
Adiwardana, D., Luong, M.-T., So, D. R., Hall, J., Fiedel, N., Thoppilan, R., Yang, Z., Kulshreshtha, A., Nemade, G., Lu, Y., et al · 2020
Closest in time.
Co-evolution of language and agents in referential games
Dagan, G., Hupkes, D., and Bruni, E · 2020
Closest in time.
Revisiting self-training for neural sequence generation
He, J., Gu, J., Shen, J., and Ranzato, M · 2020
Closest in time.
Multi-agent communication meets natural language: Synergies between functional and structural language learning
Lazaridou, A., Potapenko, A., and Tieleman, O · 2020
Closest in time.
Compositional languages emerge in a neural iterated learning model
Ren, Y., Guo, S., Labeau, M., Cohen, S. B., and Kirby, S · 2020
Closest in time.