Fetching the paper…
Reading the bibliography…
Automatic open-domain dialogue evaluation is a crucial component of dialogue systems.
Roberta: A robustly optimized bert pretraining approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
End-to-End evaluation in VERBMOBIL I
Nübel, R. 1997 · 1997
Earlier work this paper cites.
Towards a human-like open-domain chatbot
Adiwardana, D.; Luong, M.-T.; So, D. R.; Hall, J.; Fiedel, N.; Thoppilan, R.; Yang, Z.; Kulshreshtha, A.; Nemade, G.; Lu, Y.; et al. 2020 · 2001
Earlier work this paper cites.
Bleu: A method for automatic evaluation of machine translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Banerjee, S.; and Lavie, A. 2005 · 2005
Earlier work this paper cites.
Chameleons in imagined conversations: a new approach to understanding coordination of linguistic style in dialogs
Danescu-Niculescu-Mizil, C.; and Lee, L. 2011 · 2011
Earlier work this paper cites.
An optimal assessment of natural language student input using word-to-word similarity metrics
Rus, V.; and Lintean, M. 2012 · 2012
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Mikolov, T.; Chen, K.; Corrado, G.; and Dean, J. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P.; and Ba, J. 2015 · 2015
Earlier work this paper cites.
A Neural Network Approach to Context-Sensitive Generation of Conversational Responses
Sordoni, A.; Galley, M.; Auli, M.; Brockett, C.; Ji, Y.; Mitchell, M.; Nie, J.-Y.; Gao, J.; and Dolan, W. B. 2015 · 2015
Earlier work this paper cites.
Towards universal paraphrastic sentence embeddings
Wieting, J.; Bansal, M.; Gimpel, K.; and Livescu, K. 2016 · 2016
Earlier work this paper cites.
End-to-end Conversation Modeling Track in DSTC6
Hori, C.; and Hori, T. 2017 · 2017
Earlier work this paper cites.
Dailydialog: A manually labelled multi-turn dialogue dataset
Li, Y.; Su, H.; Shen, X.; Li, W.; Cao, Z.; and Niu, S. 2017 · 2017
Cited alongside, same era.
Conceptnet 5.5: An open multilingual graph of general knowledge
Speer, R.; Chin, J.; and Havasi, C. 2017 · 2017
Cited alongside, same era.
Towards implicit content-introducing for generative short-text conversation systems
Yao, L.; Zhang, Y.; Feng, Y.; Zhao, D.; and Yan, R. 2017 · 2017
Cited alongside, same era.
Learning discourse-level diversity for neural dialog models using conditional variational autoencoders
Zhao, T.; Zhao, R.; and Eskenazi, M. 2017 · 2017
Cited alongside, same era.
Ruber: An unsupervised method for automatic evaluation of open-domain dialog systems
Tao, C.; Mou, L.; Zhao, D.; and Yan, R. 2018 · 2018
Cited alongside, same era.
Graph attention networks
Veličković, P.; Cucurull, G.; Casanova, A.; Romero, A.; Liò, P.; and Bengio, Y. 2018 · 2018
What makes a good conversation? How controllable attributes affect human judgments
See, A.; Roller, S.; Kiela, D.; and Weston, J. 2019 · 2019
Later among the works it cites.
BERTScore: Evaluating text generation with BERT
Zhang, T.; Kishore, V.; Wu, F.; Weinberger, K. Q.; and Artzi, Y. 2019 · 2019
Later among the works it cites.
Predictive engagement: An efficient metric for automatic evaluation of open-domain dialogue systems
Ghazarian, S.; Weischedel, R.; Galstyan, A.; and Peng, N. 2020 · 2020
Later among the works it cites.
Grade: Automatic graph-enhanced coherence metric for evaluating open-domain dialogue systems
Huang, L.; Ye, Z.; Qin, J.; Lin, L.; and Liang, X. 2020 · 2020
Later among the works it cites.
Pone: A novel automatic evaluation metric for open-domain generative dialogue systems
Lan, T.; Mao, X.-L.; Wei, W.; Gao, X.; and Huang, H. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Retrieve and refine: Improved sequence generation models for dialogue
Weston, J.; Dinan, E.; and Miller, A. 2018 · 2018
Cited alongside, same era.
Learning to control the specificity in neural response generation
Zhang, R.; Guo, J.; Fan, Y.; Lan, Y.; Xu, J.; and Cheng, X. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Cited alongside, same era.
Grounded response generation task at dstc7
Galley, M.; Brockett, C.; Gao, X.; Gao, J.; and Dolan, B. 2019 · 2019
Cited alongside, same era.
Better automatic evaluation of open-domain dialogue systems with contextualized embeddings
Ghazarian, S.; Wei, J.; Galstyan, A.; and Peng, N. 2019 · 2019
Cited alongside, same era.
Investigating evaluation of open-domain dialogue systems with human generated multiple references
Gupta, P.; Mehri, S.; Zhao, T.; Pavel, A.; Eskenazi, M.; and Bigham, J. P. 2019 · 2019
Cited alongside, same era.
Merdivan, E.; Singh, D.; Hanke, S.; Kropf, J.; Holzinger, A.; and Geist, M. 2020 · 2020
Later among the works it cites.
Deconstruct to reconstruct a configurable evaluation metric for open-domain dialogue systems
Phy, V.; Zhao, Y.; and Aizawa, A. 2020 · 2020
Later among the works it cites.
Learning an unreferenced metric for online dialogue evaluation
Sinha, K.; Parthasarathi, P.; Wang, J.; Lowe, R.; Hamilton, W. L.; and Pineau, J. 2020 · 2020
Later among the works it cites.
Designing precise and robust dialogue response evaluators
Zhao, T.; Lala, D.; and Kawahara, T. 2020 · 2020
Later among the works it cites.
Automatic evaluation and moderation of open-domain dialogue systems
Chen, Z.; Sadoc, J.; D’Haro, L. F.; Banchs, R.; and Rudnicky, A. 2021 · 2021
Later among the works it cites.
SimCSE: Simple Contrastive Learning of Sentence Embeddings
Gao, T.; Yao, X.; and Chen, D. 2021 · 2021
Later among the works it cites.
Deep AM-FM: Toolkit for automatic dialogue evaluation
Zhang, C.; D’Haro, L. F.; Banchs, R. E.; Friedrichs, T.; and Li, H. 2021 · 2021
Later among the works it cites.