Fetching the paper…
Reading the bibliography…
The development of Open-Domain Dialogue Systems (ODS)is a trending topic due to the large number of research challenges, large societal and business impact, and advances in the underlying technology.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
DialoGPT: Large-scale generative pre-training for conversational response generation
Zhang, Y.; Sun, S.; Galley, M.; Chen, Y.-C.; Brockett, C.; Gao, X.; Gao, J.; Liu, J.; and Dolan, B. 2019 · 1911
Earlier work this paper cites.
Towards a human-like open-domain chatbot
Adiwardana, D.; Luong, M.-T.; So, D. R.; Hall, J.; Fiedel, N.; Thoppilan, R.; Yang, Z.; Kulshreshtha, A.; Nemade, G.; Lu, Y.; et al. 2020 · 2001
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
Colbert: Using bert sentence embedding for humor detection
Annamoradnejad, I.; and Zoghi, G. 2020 · 2004
Earlier work this paper cites.
ROUGE: A Package for Automatic Evaluation of Summaries
Lin, C.-Y. 2004 · 2004
Earlier work this paper cites.
Can you put it all together: Evaluating conversational agents’ ability to blend skills
Smith, E. M.; Williamson, M.; Shuster, K.; Weston, J.; and Boureau, Y.-L. 2020 · 2004
Earlier work this paper cites.
Language models are few-shot learners
Brown, T. B.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 2005
Earlier work this paper cites.
An Evaluation Protocol for Generative Conversational Systems
Lee, S.; Lim, H.; and Sedoc, J. 2020 · 2010
Earlier work this paper cites.
Recipes for safety in open-domain chatbots
Xu, J.; Ju, D.; Li, M.; Boureau, Y.-L.; Weston, J.; and Dinan, E. 2020 · 2010
Earlier work this paper cites.
Danescu-Niculescu-Mizil, C.; and Lee, L. 2011 · 2011
Earlier work this paper cites.
Movie-DiC: a movie dialogue corpus for research and development
Banchs, R. E. 2012 · 2012
Earlier work this paper cites.
Vinyals, O.; and Le, Q. 2015 · 2015
Earlier work this paper cites.
Liu, C.-W.; Lowe, R.; Serban, I. V.; Noseworthy, M.; Charlin, L.; and Pineau, J. 2016 · 2016
Earlier work this paper cites.
End-to-end conversation modeling track in DSTC6
Hori, C.; and Hori, T. 2017 · 2017
Earlier work this paper cites.
Emotionlines: An emotion corpus of multi-party conversations
Chen, S.-Y.; Hsu, C.-C.; Kuo, C.-C.; Ku, L.-W.; et al. 2018 · 2018
Earlier work this paper cites.
A Call for Clarity in Reporting BLEU Scores
Post, M. 2018 · 2018
Earlier work this paper cites.
Towards empathetic open-domain conversation models: A new benchmark and dataset
Rashkin, H.; Smith, E. M.; Li, M.; and Boureau, Y.-L. 2018 · 2018
Cited alongside, same era.
Carer: Contextualized affect representations for emotion recognition
Saravia, E.; Liu, H.-C. T.; Huang, Y.-H.; Wu, J.; and Chen, Y.-S. 2018 · 2018
Cited alongside, same era.
Personalizing dialogue agents: I have a dog, do you have pets too?
Zhang, S.; Dinan, E.; Urbanek, J.; Szlam, A.; Kiela, D.; and Weston, J. 2018 · 2018
Cited alongside, same era.
Grounded Response Generation Task at DSTC7
Galley, M.; Brockett, C.; Gao, X.; Gao, J.; and Dolan, B. 2019 · 2019
Cited alongside, same era.
Topical-Chat: Towards Knowledge-Grounded Open-Domain Conversations
Gopalakrishnan, K.; Hedayatnia, B.; Chen, Q.; Gottardi, A.; Kwatra, S.; Venkatesh, A.; Gabriel, R.; Hakkani-Tür, D.; and AI, A. A. 2019 · 2019
Cited alongside, same era.
Towards Holistic and Automatic Evaluation of Open-Domain Dialogue Generation
Pang, B.; Nijkamp, E.; Han, W.; Zhou, L.; Liu, Y.; and Tu, K. 2020 · 2020
Later among the works it cites.
Deconstruct to Reconstruct a Configurable Evaluation Metric for Open-Domain Dialogue Systems
Phy, V.; Zhao, Y.; and Aizawa, A. 2020 · 2020
Later among the works it cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Raffel, C.; Shazeer, N.; Roberts, A.; Lee, K.; Narang, S.; Matena, M.; Zhou, Y.; Li, W.; and Liu, P. J. 2020 · 2020
Later among the works it cites.
Improving Dialog Evaluation with a Multi-reference Adversarial Dataset and Large Scale Pretraining
Sai, A. B.; Mohankumar, A. K.; Arora, S.; and Khapra, M. M. 2020 · 2020
Later among the works it cites.
BLEURT: Learning Robust Metrics for Text Generation
Sellam, T.; Das, D.; and Parikh, A. P. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Investigating Evaluation of Open-Domain Dialogue Systems With Human Generated Multiple References
Gupta, P.; Mehri, S.; Zhao, T.; Pavel, A.; Eskenazi, M.; and Bigham, J. 2019 · 2019
Cited alongside, same era.
Subjective annotation and evaluation of three different chatbots WOCHAT: shared task report
Kong-Vega, N.; Shen, M.; Wang, M.; and D’Haro, L. F. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Radford, A.; Wu, J.; Child, R.; Luan, D.; Amodei, D.; Sutskever, I.; et al. 2019 · 2019
Cited alongside, same era.
Towards Empathetic Open-domain Conversation Models: A New Benchmark and Dataset
Rashkin, H.; Smith, E. M.; Li, M.; and Boureau, Y.-L. 2019 · 2019
Cited alongside, same era.
ChatEval: A Tool for Chatbot Evaluation
Sedoc, J.; Ippolito, D.; Kirubarajan, A.; Thirani, J.; Ungar, L.; and Callison-Burch, C. 2019 · 2019
Cited alongside, same era.
What makes a good conversation? How controllable attributes affect human judgments
See, A.; Roller, S.; Kiela, D.; and Weston, J. 2019 · 2019
Cited alongside, same era.
The second conversational intelligence challenge (convai2)
Dinan, E.; Logacheva, V.; Malykh, V.; Miller, A.; Shuster, K.; Urbanek, J.; Kiela, D.; Szlam, A.; Serban, I.; Lowe, R.; et al. 2020 · 2020
Cited alongside, same era.
Zhang*, T.; Kishore*, V.; Wu*, F.; Weinberger, K. Q.; and Artzi, Y. 2020 · 2020
Later among the works it cites.
Designing Precise and Robust Dialogue Response Evaluators
Zhao, T.; Lala, D.; and Kawahara, T. 2020 · 2020
Later among the works it cites.
Characterizing Social Spambots by their Human Traits
Giorgi, S.; Ungar, L.; and Schwartz, H. A. 2021 · 2021
Closest in time.
Internet-augmented dialogue generation
Komeili, M.; Shuster, K.; and Weston, J. 2021 · 2021
Closest in time.
Agreeing to Disagree: Annotating Offensive Language Datasets with Annotators’ Disagreement
Leonardelli, E.; Menini, S.; Aprosio, A. P.; Guerini, M.; and Tonelli, S. 2021 · 2021
Closest in time.
Genuine 2 : An open domain chatbot based on generative models
Rodríguez-Cantelar, M.; de la Cal, D.; Estecha, M.; Gutiérrez, A. G.; Martín, D.; Milara, N. R. N.; Jiménez, R. M.; and D’Haro, L. F. 2021 · 2021
Closest in time.
Recipes for Building an Open-Domain Chatbot
Roller, S.; Dinan, E.; Goyal, N.; Ju, D.; Williamson, M.; Liu, Y.; Xu, J.; Ott, M.; Smith, E. M.; Boureau, Y.-L.; and Weston, J. 2021 · 2021
Closest in time.
Beyond goldfish memory: Long-term open-domain conversation
Xu, J.; Szlam, A.; and Weston, J. 2021 · 2021
Closest in time.
Towards Quantifiable Dialogue Coherence Evaluation
Ye, Z.; Lu, L.; Huang, L.; Lin, L.; and Liang, X. 2021 · 2021
Closest in time.
A Comprehensive Assessment of Dialog Evaluation Metrics
Yeh, Y.-T.; Eskenazi, M.; and Mehri, S. 2021 · 2021
Closest in time.
DynaEval: Unifying Turn and Dialogue Level Evaluation
Zhang, C.; Chen, Y.; D’Haro, L. F.; Zhang, Y.; Friedrichs, T.; Lee, G.; and Li, H. 2021a · 2021
Closest in time.