Fetching the paper…
Reading the bibliography…
Conversational question answering aims to provide natural-language answers to users in information-seeking conversations.
Better automatic evaluation of open-domain dialogue systems with contextualized embeddings
Sarik Ghazarian, Johnny Tian-Zheng Wei, Aram Galstyan, and Nanyun Peng. 2019 · 1904
Earlier work this paper cites.
Machine translation evaluation with bert regressor
Hiroki Shimanaka, Tomoyuki Kajiwara, and Mamoru Komachi. 2019 · 1907
Earlier work this paper cites.
Measuring nominal scale agreement among many raters
Joseph L Fleiss. 1971 · 1971
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Chia-Wei Liu, Ryan Lowe, Iulian V Serban, Michael Noseworthy, Laurent Charlin, and Joelle Pineau. 2016 · 2016
Earlier work this paper cites.
SQuAD: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Earlier work this paper cites.
Visual dialog
Abhishek Das, Satwik Kottur, Khushi Gupta, Avi Singh, Deshraj Yadav, José M. F. Moura, Devi Parikh, and Dhruv Batra. 2017 · 2017
Earlier work this paper cites.
Towards an automatic Turing test: Learning to evaluate dialogue responses
Ryan Lowe, Michael Noseworthy, Iulian Vlad Serban, Nicolas Angelard-Gontier, Yoshua Bengio, and Joelle Pineau. 2017 · 2017
Earlier work this paper cites.
ParlAI: A dialog research software platform
Alexander Miller, Will Feng, Dhruv Batra, Antoine Bordes, Adam Fisch, Jiasen Lu, Devi Parikh, and Jason Weston. 2017 · 2017
Earlier work this paper cites.
QuAC: Question answering in context
Eunsol Choi, He He, Mohit Iyyer, Mark Yatskar, Wen-tau Yih, Yejin Choi, Percy Liang, and Luke Zettlemoyer. 2018 · 2018
Earlier work this paper cites.
AllenNLP: A deep semantic natural language processing platform
Matt Gardner, Joel Grus, Mark Neumann, Oyvind Tafjord, Pradeep Dasigi, Nelson F. Liu, Matthew Peters, Michael Schmitz, and Luke Zettlemoyer. 2018 · 2018
Earlier work this paper cites.
Dialog-to-action: Conversational question answering over a large-scale knowledge base
Daya Guo, Duyu Tang, Nan Duan, Ming Zhou, and Jian Yin. 2018 · 2018
Cited alongside, same era.
Higher-order coreference resolution with coarse-to-fine inference
Kenton Lee, Luheng He, and Luke Zettlemoyer. 2018 · 2018
Cited alongside, same era.
Amrita Saha, Vardaan Pahuja, Mitesh M. Khapra, Karthik Sankaranarayanan, and Sarath Chandar. 2018 · 2018
Cited alongside, same era.
Ruber: An unsupervised method for automatic evaluation of open-domain dialog systems
Chongyang Tao, Lili Mou, Dongyan Zhao, and Rui Yan. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
GRADE: Automatic graph-enhanced coherence metric for evaluating open-domain dialogue systems
Lishan Huang, Zheng Ye, Jinghui Qin, Liang Lin, and Xiaodan Liang. 2020 · 2020
Later among the works it cites.
USR: An unsupervised and reference free evaluation metric for dialog generation
Shikib Mehri and Maxine Eskenazi. 2020 · 2020
Later among the works it cites.
Open-Retrieval Conversational Question Answering , page 539–548. Association for Computing Machinery, New York, NY, USA
Chen Qu, Liu Yang, Cen Chen, Minghui Qiu, W. Bruce Croft, and Mohit Iyyer. 2020 · 2020
Later among the works it cites.
Improving dialog evaluation with a multi-reference adversarial dataset and large scale pretraining
Ananya B. Sai, Akash Kumar Mohankumar, Siddhartha Arora, and Mitesh M. Khapra. 2020 · 2020
Later among the works it cites.
Topiocqa: Open-domain conversational question answering with topic switching
Vaibhav Adlakha, Shehzaad Dhuliawala, Kaheer Suleman, Harm de Vries, and Siva Reddy. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Can you unpack that? learning to rewrite questions-in-context
Ahmed Elgohary, Denis Peskov, and Jordan Boyd-Graber. 2019 · 2019
Cited alongside, same era.
Investigating evaluation of open-domain dialogue systems with human generated multiple references
Prakhar Gupta, Shikib Mehri, Tiancheng Zhao, Amy Pavel, Maxine Eskenazi, and Jeffrey Bigham. 2019 · 2019
Cited alongside, same era.
A simple but effective method to incorporate multi-turn context with BERT for conversational machine comprehension
Yasuhito Ohsugi, Itsumi Saito, Kyosuke Nishida, Hisako Asano, and Junji Tomita. 2019 · 2019
Cited alongside, same era.
CoQA: A conversational question answering challenge
Siva Reddy, Danqi Chen, and Christopher D. Manning. 2019 · 2019
Cited alongside, same era.
Re-evaluating adem: A deeper look at scoring dialogue responses
Ananya B Sai, Mithun Das Gupta, Mitesh M Khapra, and Mukundhan Srinivasan. 2019 · 2019
Cited alongside, same era.
DoQA - accessing domain-specific FAQs via conversational QA
Jon Ander Campos, Arantxa Otegi, Aitor Soroa, Jan Deriu, Mark Cieliebak, and Eneko Agirre. 2020 · 2020
Cited alongside, same era.
Yu Chen, Lingfei Wu, and Mohammed J Zaki. 2020 · 2020
Cited alongside, same era.
Closest in time.
Open-domain question answering goes conversational via question rewriting
Raviteja Anantha, Svitlana Vakulenko, Zhucheng Tu, Shayne Longpre, Stephen Pulman, and Srinivas Chappidi. 2021 · 2021
Closest in time.
Learn to resolve conversational dependency: A consistency training framework for conversational question answering
Gangwoo Kim, Hyunjae Kim, Jungsoo Park, and Jaewoo Kang. 2021 · 2021
Closest in time.
Towards a more robust evaluation for conversational question answering
Wissam Siblini, Baris Sayil, and Yacine Kessaci. 2021 · 2021
Closest in time.
Towards quantifiable dialogue coherence evaluation
Zheng Ye, Liucun Lu, Lishan Huang, Liang Lin, and Xiaodan Liang. 2021 · 2021
Closest in time.
Do not let the history haunt you: Mitigating compounding errors in conversational question answering
Angrosh Mandya, James O’ Neill, Danushka Bollegala, and Frans Coenen. 2020 · 2025
Closest in time.
Interpretation of natural language rules in conversational machine reading
Marzieh Saeidi, Max Bartolo, Patrick Lewis, Sameer Singh, Tim Rocktäschel, Mike Sheldon, Guillaume Bouchard, and Sebastian Riedel. 2018 · 2097
Closest in time.