Fetching the paper…
Reading the bibliography…
Next generation task-oriented dialog systems need to understand conversational contexts with their perceived surroundings, to effectively help users in the real-world multimodal environment.
Clevr-dialog: A diagnostic dataset for multi-round reasoning in visual dialog
Satwik Kottur, José MF Moura, Devi Parikh, Dhruv Batra, and Marcus Rohrbach. 2019 · 1903
Earlier work this paper cites.
Multiwoz 2.1: Multi-domain dialogue state corrections and state tracking baselines
Mihail Eric, Rahul Goel, Shachi Paul, Adarsh Kumar, Abhishek Sethi, Peter Ku, Anuj Kumar Goyal, Sanchit Agarwal, Shuyag Gao, and Dilek Hakkani-Tur. 2019 · 1907
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
A simple language model for task-oriented dialogue
Ehsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz, and Richard Socher. 2020 · 2005
Earlier work this paper cites.
Soloist: Few-shot task-oriented dialog with a single pre-trained auto-regressive model
Baolin Peng, Chunyuan Li, Jinchao Li, Shahin Shayandeh, Lars Liden, and Jianfeng Gao. 2020 · 2005
Earlier work this paper cites.
Situated and interactive multimodal conversations
Seungwhan Moon, Satwik Kottur, Paul A Crook, Ankita De, Shivani Poddar, Theodore Levin, David Whitney, Daniel Difranco, Ahmad Beirami, Eunjoon Cho, Rajen Subba, and Alborz Geramifard. 2020 · 2006
Earlier work this paper cites.
Agenda-based user simulation for bootstrapping a pomdp dialogue system
Jost Schatzmann, Blaise Thomson, Karl Weilhammer, Hui Ye, and Steve Young. 2007 · 2007
Earlier work this paper cites.
Overview of the ninth dialog system technology challenge: Dstc9
Chulaka Gunasekara, Seokhwan Kim, Luis Fernando D’Haro, Abhinav Rastogi, Yun-Nung Chen, Mihail Eric, Behnam Hedayatnia, Karthik Gopalakrishnan, Yang Liu, Chao-Wei Huang, et al. 2020 · 2011
Earlier work this paper cites.
Example-based synthesis of 3d object arrangements
Matthew Fisher, Daniel Ritchie, Manolis Savva, Thomas Funkhouser, and Pat Hanrahan. 2012 · 2012
Earlier work this paper cites.
The second dialog state tracking challenge
Matthew Henderson, Blaise Thomson, and Jason D Williams. 2014 · 2014
Earlier work this paper cites.
VQA: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Activity-centric scene synthesis for functional 3d scene modeling
Matthew Fisher, Manolis Savva, Yangyan Li, Pat Hanrahan, and Matthias Nießner. 2015 · 2015
Cited alongside, same era.
Visual dialog
Abhishek Das, Satwik Kottur, Khushi Gupta, Avi Singh, Deshraj Yadav, José MF Moura, Devi Parikh, and Dhruv Batra. 2017 · 2017
Cited alongside, same era.
Guesswhat?! visual object discovery through multi-modal dialogue
Harm de Vries, Florian Strub, Sarath Chandar, Olivier Pietquin, Hugo Larochelle, and Aaron Courville. 2017 · 2017
Cited alongside, same era.
MultiWOZ - a large-scale multi-domain wizard-of-Oz dataset for task-oriented dialogue modelling
Paweł Budzianowski, Tsung-Hsien Wen, Bo-Hsiang Tseng, Iñigo Casanueva, Stefan Ultes, Osman Ramadan, and Milica Gašić. 2018 · 2018
Cited alongside, same era.
Audio visual scene-aware dialog track in dstc8
Multimodal transformer networks for end-to-end video-grounded dialogue systems
Hung Le, Doyen Sahoo, Nancy Chen, and Steven Hoi. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Towards scalable multi-domain conversational agents: The schema-guided dialogue dataset
Abhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta, and Pranav Khaitan. 2019 · 2019
Later among the works it cites.
Situated interactive multimodal conversations (simmc) track at dstc9
Paul A. Crook, Satwik Kottur, Seungwhan Moon, Ahmad Beirami, Eunjoon Cho, Rajen Subba, and Alborz Geramifard. 2021 · 2021
Closest in time.
Joint generation and bi-encoder for situated interactive multimodal conversations
Xin Huang, Chor Seng Tan, Yan Bin Ng, Wei Shi, Kheng Hui Yeo, Ridong Jiang, and Jung Jae Kim. 2021 · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chiori Hori, Anoop Cherian, Tim K. Marks, and Florian Metze. 2018 · 2018
Cited alongside, same era.
Building a conversational agent overnight with dialogue self-play
Pararth Shah, Dilek Hakkani-Tür, Gokhan Tür, Abhinav Rastogi, Ankur Bapna, Neha Nayak, and Larry Heck. 2018 · 2018
Cited alongside, same era.
Talk the walk: Navigating new york city through grounded dialogue
Harm de Vries, Kurt Shuster, Dhruv Batra, Devi Parikh, Jason Weston, and Douwe Kiela. 2018 · 2018
Cited alongside, same era.
Bert-dst: Scalable end-to-end dialogue state tracking with bidirectional encoder representations from transformer
Guan-Lin Chao and Ian Lane. 2019 · 2019
Cited alongside, same era.
Dialog state tracking: A neural reading comprehension approach
Shuyang Gao, Sanchit Agarwal Abhishek Seth and, Tagyoung Chun, and Dilek Hakkani-Ture. 2019 · 2019
Cited alongside, same era.
Tom : End-to-end task-oriented multimodal dialog system with gpt-2
Younghoon Jeong, Se Jin Lee, Youngjoong Ko, and Jungyun Seo. 2021 · 2021
Closest in time.
Improving multimodal api prediction via adding dialog state and various multimodal gates
Byoungjae Kim, Inkwon Lee, Yeonseok Jeong, Ko Youngjoong, Myoung-Wan Koo, and Jungyun Seo. 2021 · 2021
Closest in time.
Multi-task learning for situated multi-domain end-to-end dialogue systems
Po-Nien Kung, Tse-Hsuan Yang, Chung-Cheng Chang, Hsin-Kai Hsu, Yu-Jia Liou, and Yun-Nung Chen. 2021 · 2021
Closest in time.
A response retrieval approach for dialogue using a multi-attentive transformer
Matteo Antonio Senese, Giuseppe Rizzo, Alberto Benincasa, and Barbara Caputo. 2021 · 2021
Closest in time.