Fetching the paper…
Reading the bibliography…
We consider grounding open domain dialogues with images.
Neural Response Generation with Meta-Words
Xu, C.; Wu, W.; Tao, C.; Hu, H.; Schuerman, M.; and Wang, Y. 2019 · 1906
Earlier work this paper cites.
Thomason, J.; Murray, M.; Cakmak, M.; and Zettlemoyer, L. 2019 · 1907
Earlier work this paper cites.
The equivalence of weighted kappa and the intraclass correlation coefficient as measures of reliability
Fleiss, J. L.; and Cohen, J. 1973 · 1973
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
Low-Resource Knowledge-Grounded Dialogue Generation
Zhao, X.; Wu, W.; Tao, C.; Xu, C.; Zhao, D.; and Yan, R. 2020 · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y. 2004 · 2004
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J.; Dong, W.; Socher, R.; Li, L.-J.; Li, K.; and Fei-Fei, L. 2009 · 2009
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P.; and Welling, M. 2013 · 2013
Earlier work this paper cites.
On the Properties of Neural Machine Translation: Encoder–Decoder Approaches
Cho, K.; van Merriënboer, B.; Bahdanau, D.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, I.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks
Sutskever, I.; Vinyals, O.; and Le, Q. V. 2014 · 2014
Earlier work this paper cites.
Vqa: Visual question answering
Antol, S.; Agrawal, A.; Lu, J.; Mitchell, M.; Batra, D.; Lawrence Zitnick, C.; and Parikh, D. 2015 · 2015
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Bahdanau, D.; Cho, K.; and Bengio, Y. 2015 · 2015
Earlier work this paper cites.
Adam: A method for Stochastic Optimization
Kingma, D. P.; and Ba, J. 2015 · 2015
Earlier work this paper cites.
Makhzani, A.; Shlens, J.; Jaitly, N.; Goodfellow, I.; and Frey, B. 2015 · 2015
Earlier work this paper cites.
Neural Responding Machine for Short-Text Conversation
Shang, L.; Lu, Z.; and Li, H. 2015 · 2015
Earlier work this paper cites.
Learning Structured Output Representation Using Deep Conditional Generative Models
Sohn, K.; Yan, X.; and Lee, H. 2015 · 2015
Cited alongside, same era.
Vinyals, O.; and Le, Q. 2015 · 2015
Cited alongside, same era.
Show and tell: A neural image caption generator
Vinyals, O.; Toshev, A.; Bengio, S.; and Erhan, D. 2015 · 2015
Cited alongside, same era.
A Diversity-Promoting Objective Function for Neural Conversation Models
Li, J.; Galley, M.; Brockett, C.; Gao, J.; and Dolan, B. 2015 · 2016
Cited alongside, same era.
A Persona-Based Neural Conversation Model
Li, J.; Galley, M.; Brockett, C.; Spithourakis, G.; Gao, J.; and Dolan, B. 2016 · 2016
Cited alongside, same era.
Building End-To-End Dialogue Systems Using Generative Hierarchical Neural Network Models
Talk the Walk: Navigating New York City through Grounded Dialogue
de Vries, H.; Shuster, K.; Batra, D.; Parikh, D.; Weston, J.; and Kiela, D. 2018 · 2018
Later among the works it cites.
Wizard of wikipedia: Knowledge-powered conversational agents
Dinan, E.; Roller, S.; Shuster, K.; Fan, A.; Auli, M.; and Weston, J. 2018 · 2018
Later among the works it cites.
Augmenting Neural Response Generation with Context-Aware Topical Attention
Dziri, N.; Kamalloo, E.; Mathewson, K. W.; and Zaiane, O. 2018 · 2018
Later among the works it cites.
Emotional dialogue generation using image-grounded language models
Huber, B.; McDuff, D.; Brockett, C.; Galley, M.; and Dolan, B. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Serban, I. V.; Sordoni, A.; Bengio, Y.; Courville, A. C.; and Pineau, J. 2016 · 2016
Cited alongside, same era.
Rethinking the inception architecture for computer vision
Szegedy, C.; Vanhoucke, V.; Ioffe, S.; Shlens, J.; and Wojna, Z. 2016 · 2016
Cited alongside, same era.
Visual dialog
Das, A.; Kottur, S.; Gupta, K.; Singh, A.; Yadav, D.; Moura, J. M.; Parikh, D.; and Batra, D. 2017 · 2017
Cited alongside, same era.
Image-Grounded Conversations: Multimodal Context for Natural Question and Response Generation
Mostafazadeh, N.; Brockett, C.; Dolan, B.; Galley, M.; Gao, J.; Spithourakis, G.; and Vanderwende, L. 2017 · 2017
Cited alongside, same era.
A Hierarchical Latent Variable Encoder-Decoder Model for Generating Dialogues
Serban, I. V.; Sordoni, A.; Lowe, R.; Charlin, L.; Pineau, J.; Courville, A. C.; and Bengio, Y. 2017 · 2017
Cited alongside, same era.
Attention is All you Need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, L. u.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
Topic Aware Neural Response Generation
Xing, C.; Wu, W.; Wu, Y.; Liu, J.; Huang, Y.; Zhou, M.; and Ma, W.-Y. 2017 · 2017
Cited alongside, same era.
Ram, A.; Prasad, R.; Khatri, C.; Venkatesh, A.; Gabriel, R.; Liu, Q.; Nunn, J.; Hedayatnia, B.; Cheng, M.; Nagar, A.; et al. 2018 · 2018
Later among the works it cites.
Chatpainter: Improving text to image generation using dialogue
Sharma, S.; Suhubdy, D.; Michalski, V.; Kahou, S. E.; and Bengio, Y. 2018 · 2018
Later among the works it cites.
From Eliza to XiaoIce: Challenges and Opportunities with Social Chatbots
Shum, H.; He, X.; and Li, D. 2018 · 2018
Later among the works it cites.
Hierarchical Recurrent Attention Network for Response Generation
Xing, C.; Wu, W.; Wu, Y.; Zhou, M.; Huang, Y.; and Ma, W.-Y. 2018 · 2018
Later among the works it cites.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
Xu, T.; Zhang, P.; Huang, Q.; Zhang, H.; Gan, Z.; Huang, X.; and He, X. 2018 · 2018
Later among the works it cites.
A Dataset for Document Grounded Conversations
Zhou, K.; Prabhumoye, S.; and Black, A. W. 2018 · 2018
Later among the works it cites.
End-to-end audio visual scene-aware dialog using multimodal attention-based video features
Hori, C.; Alamri, H.; Wang, J.; Wichern, G.; Hori, T.; Cherian, A.; Marks, T. K.; Cartillier, V.; Lopes, R. G.; Das, A.; et al. 2019 · 2019
Later among the works it cites.
Multimodal Transformer Networks for End-to-End Video-Grounded Dialogue Systems
Le, H.; Sahoo, D.; Chen, N.; and Hoi, S. 2019 · 2019
Later among the works it cites.
MirrorGAN: Learning Text-to-image Generation by Redescription
Qiao, T.; Zhang, J.; Xu, D.; and Tao, D. 2019 · 2019
Later among the works it cites.
ReCoSa: Detecting the Relevant Contexts with Self-Attention for Multi-turn Dialogue Generation
Zhang, H.; Lan, Y.; Pang, L.; Guo, J.; and Cheng, X. 2019 · 2019
Later among the works it cites.
Image-Chat: Engaging Grounded Conversations
Shuster, K.; Humeau, S.; Bordes, A.; and Weston, J. 2020 · 2020
Closest in time.