Fetching the paper…
Reading the bibliography…
In this paper we survey the methods and concepts developed for the evaluation of dialogue systems.
Gupta P, Mehri S, Zhao T, Pavel A, Eskenazi M, Bigham JP (2019) Investigating evaluation of open-domain dialogue systems with human generated multiple references. 1907.10568
1907
Earlier work this paper cites.
Liu Y, Ott M, Goyal N, Du J, Joshi M, Chen D, Levy O, Lewis M, Zettlemoyer L, Stoyanov V (2019) Roberta: A robustly optimized bert pretraining approach. 1907.11692
1907
Earlier work this paper cites.
Ju Y, Zhao F, Chen S, Zheng B, Yang X, Liu Y (2019) Technical report on conversational question answering. 1909.10772
1909
Earlier work this paper cites.
Lan Z, Chen M, Goodman S, Gimpel K, Sharma P, Soricut R (2019) Albert: A lite bert for self-supervised learning of language representations. 1909.11942
1909
Earlier work this paper cites.
Turing AM (1950) Computing Machinery and Intelligence. Mind LIX(236):433–460, DOI 10.1093/mind/LIX.236.433
1950
Earlier work this paper cites.
Austin JL (1962) How to do things with words. William James Lectures, Oxford University Press
1962
Earlier work this paper cites.
Levenshtein VI (1966) Binary codes capable of correcting deletions, insertions and reversals. Soviet Physics Doklady 10(8):707–710, doklady Akademii Nauk SSSR, V163 No4 845-848 1965
1965
Earlier work this paper cites.
Weizenbaum J (1966) ELIZA - a Computer Program for the Study of Natural Language Communication Between Man and Machine. Communications of the ACM 9(1):36–45, URL http://doi.acm.org/10.1145/365153.365168
1966
Earlier work this paper cites.
Searle JR (1969) Speech Acts: An Essay in the Philosophy of Language. Cambridge University Press, Cambridge, London
1969
Earlier work this paper cites.
Fleiss JL (1971) Measuring nominal scale agreement among many raters. Psychological bulletin 76(5):378–382, DOI http://dx.doi.org/10.1037/h0031619
1971
Earlier work this paper cites.
Searle JR (1975) Indirect speech acts. In: Cole P, Morgan J (eds) Syntax and Semantics 3: Speech Acts, Academic Press, New York, pp 59–82
1975
Earlier work this paper cites.
Colby KM (1981) Modeling a paranoid mind. Behavioral and Brain Sciences 4(4):515–534
1981
Earlier work this paper cites.
Hirschman L, Dahl DA, McKay DP, Norton LM, Linebarger MC (1990) Beyond class A: A proposal for automatic evaluation of discourse. In: In Proceedings of the Speech and Natural Language Workshop, Hidden Valley, Pennsylvania, USA, HLT, pp 109–113
1990
Earlier work this paper cites.
Godfrey JJ, Holliman EC, McDaniel J (1992) SWITCHBOARD: telephone speech corpus for research and development. In: [Proceedings] ICASSP-92: 1992 IEEE International Conference on Acoustics, Speech, and Signal Processing, San Francisco, CA, USA, vol 1, pp 517–520, DOI 10.1109/ICASSP.1992.225858
1992
Earlier work this paper cites.
Leech GN (1993) 100 million words of english: the british national corpus (BNC). English Today 28:9–15, DOI doi:10.1017/S0266078400006854
1993
Earlier work this paper cites.
Carletta J (1996) Assessing Agreement on Classification Tasks: The Kappa Statistic. Computational Linguistics 22(2):249–254
1996
Earlier work this paper cites.
Hochreiter S, Schmidhuber J (1997) Long short-term memory. Neural Computation pp 1735–1780
1997
Earlier work this paper cites.
Walker MA, Litman DJ, Kamm CA, Abella A (1997) PARADISE: A Framework for Evaluating Spoken Dialogue Agents. In: Proceedings of the Eighth Conference on European Chapter of the Association for Computational Linguistics, Madrid, Spain, EACL ’97, pp 271–280, URL https://doi.org/10.3115/979617.979652
1997
Earlier work this paper cites.
Levin E, Pieraccini R, Eckert W (1998) Using Markov decision process for learning dialogue strategies. In: Proceedings of the 1998 IEEE International Conference on Acoustics, Speech and Signal Processing, Seattle, WA, USA, ICASSP, vol 1, pp 201–204, DOI 10.1109/ICASSP.1998.674402
1998
Earlier work this paper cites.
Cole R (1999) Tools for research and education in speech science. In: Proceedings of the International Conference of Phonetic Sciences, San Francisco, USA, pp 1277–1280
1999
Earlier work this paper cites.
Traum DR (1999) Speech Acts for Dialogue Agents, Springer Netherlands, Dordrecht, pp 169–201. URL https://doi.org/10.1007/978-94-015-9204-8_8
1999
Earlier work this paper cites.
Lamel L, Rosset S, Gauvain JL, Bennacef S, Garnier-Rizet M, Prouts B (2000) The limsi arise system. Speech Communication 31(4):339–353
2000
Earlier work this paper cites.
Singh SP, Kearns MJ, Litman DJ, Walker MA (2000) Reinforcement Learning for Spoken Dialogue Systems. In: Solla SA, Leen TK, Müller K (eds) Advances in Neural Information Processing Systems 12, MIT Press, pp 956–962, URL http://papers.nips.cc/paper/1775-reinforcement-learning-for-spoken-dialogue-systems.pdf
2000
Earlier work this paper cites.
Walker MA, Kamm CA, Litman DJ (2000) Towards developing general models of usability with PARADISE. Natural Language Engineering 6(3-4):363–377, DOI https://doi.org/10.1017/S1351324900002503
2000
Earlier work this paper cites.
Chotimongkol A, Rudnicky AI (2001) N-best speech hypotheses reordering using linear regression. In: Dalsgaard P, Lindberg B, Benner H, Tan Z (eds) EUROSPEECH 2001 Scandinavia, 7th European Conference on Speech Communication and Technology, 2nd INTERSPEECH Event, Aalborg, Denmark, September 3-7, 2001, ISCA, pp 1829–1832, URL http://www.isca-speech.org/archive/eurospeech_2001/e01_1829.html
2001
Earlier work this paper cites.
Rambow O, Bangalore S, Walker M (2001) Natural Language Generation in Dialog Systems. In: Proceedings of the first international conference on Human language technology (HLT) research, San Diego, USA, pp 67–73
2001
Earlier work this paper cites.
Papineni K, Roukos S, Ward T, Zhu WJ (2002) Bleu: a Method for Automatic Evaluation of Machine Translation. In: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics, Philadelphia, Pennsylvania, USA, ACL ’02, pp 311–318, URL http://www.aclweb.org/anthology/P02-1040
2002
Earlier work this paper cites.
Qu Y, Green N (2002) A constraint-based approach for cooperative information-seeking dialogue. In: Proceedings of the International Natural Language Generation Conference, Harriman, New York, USA, INLG, pp 136–143
2002
Earlier work this paper cites.
2002
Earlier work this paper cites.
Lin CY (2004) ROUGE: A Package for Automatic Evaluation of Summaries. In: Marie-Francine Moens SS (ed) Text Summarization Branches Out: Proceedings of the ACL-04 Workshop, Association for Computational Linguistics, Barcelona, Spain, pp 74–81, URL http://www.aclweb.org/anthology/W04-1013
2004
Earlier work this paper cites.
Möller S, Krebber J, Raake A, Smeele P, Rajman M, Melichar M, Pallotta V, Tsakou G, Kladis B, Vovos A, Hoonhout J, Schuchardt D, Fakotakis N, Ganchev T, Potamitis I (2004) INSPIRE: Evaluation of a smart-home system for infotainment management and device control. In: Proceedings of the Fourth International Conference on Language Resources and Evaluation (LREC’04), European Language Resources Association (ELRA), Lisbon, Portugal, URL http://www.lrec-conf.org/proceedings/lrec2004/pdf/12.pdf
2004
Earlier work this paper cites.
Stent A, Prasad R, Walker M (2004) Trainable Sentence Planning for Complex Information Presentation in Spoken Dialog Systems. In: Proceedings of the 42nd Annual Meeting of the Association for Computational Linguistics, Barcelona, Spain, ACL ’04, pp 79–86, URL https://www.aclweb.org/anthology/P04-1011
2004
Earlier work this paper cites.
Engel Y, Mannor S, Meir R (2005) Reinforcement Learning with Gaussian Processes. In: Proceedings of the 22nd International Conference on Machine Learning, ACM, Bonn, Germany, ICML ’05, pp 201–208
2005
Earlier work this paper cites.
McTear M, O’Neill I, Hanna P, Liu X (2005) Handling errors and determining confirmation strategies—an object-based approach. Speech Communication 45(3):249–269, URL https://www.aclweb.org/anthology/N16-1086
2005
Earlier work this paper cites.
Schatztnann J, Stuttle MN, Weilhammer K, Young S (2005) Effects of the user model on simulation-based learning of dialogue strategies. In: IEEE Workshop on Automatic Speech Recognition and Understanding, San Juan, Puerto Rico, ASRU, pp 220–225, URL https://ieeexplore.ieee.org/document/1566539
2005
Earlier work this paper cites.
Möller S, Englert R, Engelbrecht K, Hafner V, Jameson A, Oulasvirta A, Raake A, Reithinger N (2006) MeMo: towards automatic usability evaluation of spoken dialogue services by user error simulations. In: Ninth International Conference on Spoken Language Processing, INTERSPEECH - ICSLP 2006, pp 1786–1789, URL https://www.isca-speech.org/archive/interspeech_2006/i06_1131.html
2006
Earlier work this paper cites.
Paek T (2006) Reinforcement learning for spoken dialogue systems: Comparing strengths and weaknesses for practical deployment. In: Proc. Dialog-on-Dialog Workshop, Interspeech, Pittsburgh, PA, USA, URL http://www.ling.helsinki.fi/~kjokinen/ICSLP06-DoD/Programme/PaekTim.pdf
2006
Earlier work this paper cites.
Schatzmann J, Weilhammer K, Stuttle M, Young S (2006) A survey of statistical user simulation techniques for reinforcement-learning of dialogue management strategies. The Knowledge Engineering Review 21(2):97–126, URL https://search.proquest.com/docview/217517188?accountid=15920
2006
Earlier work this paper cites.
Voorhees EM (2006) Evaluating Question Answering System Performance, Springer Netherlands, Dordrecht, pp 409–430. DOI 10.1007/978-1-4020-4746-6_13
2006
Earlier work this paper cites.
Schatzmann J, Thomson B, Weilhammer K, Ye H, Young S (2007) Agenda-based User Simulation for Bootstrapping a POMDP Dialogue System. In: Human Language Technologies 2007: The Conference of the North American Chapter of the Association for Computational Linguistics; Companion Volume, Short Papers, Rochester, New York, NAACL-Short ’07, pp 149–152, URL http://dl.acm.org/citation.cfm?id=1614108.1614146
2007
Earlier work this paper cites.
van Schooten B, Rosset S, Galibert O, Max A, op den Akker R, Illouz G (2007) Handling speech input in the Ritel QA dialogue system. In: 8th Annual Conference of the International Speech Communication Association, Antwerp, Belgium, INTERSPEECH 2007, pp 126–129, URL https://www.isca-speech.org/archive/interspeech_2007/i07_0126.html
2007
Earlier work this paper cites.
Young S (2007) CUED standard dialogue acts. Report, Cambridge University, Engineering Department URL http://mi.eng.cam.ac.uk/research/dialogue/LocalDocs/dastd.pdf
2007
Earlier work this paper cites.
Young S, Schatzmann J, Weilhammer K, Ye H (2007) The Hidden Information State Approach to Dialog Management. In: IEEE International Conference on Acoustics, Speech and Signal Processing, Honolulu, HI, USA, ICASSP ’07, vol 4, pp 149–152, URL http://svr-ftp.eng.cam.ac.uk/~sjy/papers/yswy07.pdf
2007
Earlier work this paper cites.
Engelbrecht KP, Möller S, Schleicher R, Wechsung I (2008) Analysis of paradise models for individual users of a spoken dialog system. In: Electronic Speech Signal Processing, Proceedings of the 19th Conference, Frankfurt am Main, Germany, ESSV 2008, pp 86–93, URL https://d-nb.info/990359174/04
2008
Earlier work this paper cites.
Evanini K, Hunter P, Liscombe J, Suendermann D, Dayanidhi K, Pieraccini R (2008) Caller Experience: A method for evaluating dialog systems and its automatic prediction. In: 2008 IEEE Spoken Language Technology Workshop, Goa, India, pp 129–132, DOI 10.1109/SLT.2008.4777857
2008
Earlier work this paper cites.
Black AW, Eskenazi M (2009) The Spoken Dialogue Challenge. In: Proceedings of the SIGDIAL 2009 Conference: The 10th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Association for Computational Linguistics, Stroudsburg, PA, USA, SIGDIAL ’09, pp 337–340
2009
Earlier work this paper cites.
Engelbrecht KP, Gödde F, Hartard F, Ketabdar H, Möller S (2009a) Modeling User Satisfaction with Hidden Markov Model. In: Proceedings of the SIGDIAL 2009 Conference: The 10th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Association for Computational Linguistics, London, UK, SIGDIAL ’09, pp 170–177, URL http://dl.acm.org/citation.cfm?id=1708376.1708402
2009
Earlier work this paper cites.
Engelbrecht KP, Quade M, Möller S (2009b) Analysis of a New Simulation Approach to Dialog System Evaluation. Speech Communication 51(12):1234–1252, URL http://dx.doi.org/10.1016/j.specom.2009.06.007
2009
Earlier work this paper cites.
Gandhe S, Whitman N, Traum D, Artstein R (2009) An integrated authoring tool for tactical questioning dialogue systems. In: 6th IJCAI Workshop on Knowledge and Reasoning in Practical Dialogue Systems, Pasadena Conference Center, California, USA., pp 10–18
2009
Earlier work this paper cites.
Kelly D, Kantor PB, Morse EL, Scholtz J, Sun Y (2009) Questionnaires for eliciting evaluation data from users of interactive question answering systems. Natural Language Engineering 15(1):119–141
2009
Earlier work this paper cites.
Kenny PG, Parsons TD, Rizzo AA (2009) Human Computer Interaction in Virtual Standardized Patient Systems. In: Proceedings of the 13th International Conference on Human-Computer Interaction. Part IV: Interacting in Various Application Domains, Springer-Verlag, Berlin, Heidelberg, pp 514–523, URL http://dx.doi.org/10.1007/978-3-642-02583-9_56
2009
Earlier work this paper cites.
Lavie A, Denkowski MJ (2009) The Meteor Metric for Automatic Evaluation of Machine Translation. Machine Translation 23(2-3):105–115, URL http://dx.doi.org/10.1007/s10590-009-9059-4
2009
Earlier work this paper cites.
Lee C, Jung S, Kim S, Lee GG (2009) Example-based dialog modeling for practical multi-domain dialog system. Speech Communication 51(5):466–484
2009
Earlier work this paper cites.
Rieser V, Lemon O (2009) Does this list contain what you were searching for? Learning adaptive dialogue strategies for interactive question answering. Natural Language Engineering 15(1):55––72, DOI 10.1017/S1351324908004907
2009
Earlier work this paper cites.
Tiedemann J (2009) News from OPUS : A Collection of Multilingual Parallel Corpora with Tools and Interfaces. In: Recent Advances in Natural Language Processing V, vol V, John Benjamins, pp 237–248
2009
Earlier work this paper cites.
Young S, Gašić M, Keizer S, Mairesse F, Schatzmann J, Thomson B, Yu K (2010) The Hidden Information State Model: A Practical Framework for POMDP-based Spoken Dialogue Management. Computer Speech and Language 24(2):150–174, DOI 10.1016/j.csl.2009.04.001
2009
Earlier work this paper cites.
Bernardi R, Kirschner M (2010) From artificial questions to real user interaction logs: Real challenges for Interactive Question Answering systems. In: Proceedings of Workshop on Web Logs and Question Answering (WLQA’10), Valletta, Malta
2010
Earlier work this paper cites.
Hahn S, Dinarelli M, Raymond C, Lefèvre F, Lehen P, De Mori R, Moschitti A, Ney H, Riccardi G (2010) Comparing Stochastic Approaches to Spoken Language Understanding in Multiple Languages. IEEE Transactions on Audio, Speech and Language Processing (TASLP) 16:1569–1583, URL https://hal.archives-ouvertes.fr/file/index/docid/746965/filename/plugin-05639034.pdf
2010
Earlier work this paper cites.
Hara S (2010) Estimation method of user satisfaction using N-gram-based dialog history model for spoken dialog system. In: Proceedings of the Seventh International Conference on Language Resources and Evaluation, Valletta, Malta, LREC’10, pp 78–83, URL http://www.lrec-conf.org/proceedings/lrec2010/pdf/579_Paper.pdf
2010
Earlier work this paper cites.
Higashinaka R, Minami Y, Dohsaka K, Meguro T (2010) Issues in Predicting User Satisfaction Transitions in Dialogues: Individual Differences, Evaluation Criteria, and Prediction Models. In: Lee GG, Mariani J, Minker W, Nakamura S (eds) Second International Workshop on Spoken Dialogue Systems Technology: Spoken Dialogue Systems for Ambient Environments, Springer Berlin Heidelberg, Gotemba, Shizuoka, Japan, IWSDS 2010, pp 48–60
2010
Earlier work this paper cites.
Mairesse F, Gašić M, Jurčíček F, Keizer S, Thomson B, Yu K, Young S (2010) Phrase-based statistical language generation using graphical models and active learning. In: Proceedings of the 48th Annual Meeting of the Association for Computational Linguistics, Uppsala, Sweden, ACL ’10, pp 1552–1561, URL https://www.aclweb.org/anthology/P10-1157
2010
Earlier work this paper cites.
Peñas A, Magnini B, Forner P, Sutcliffe R, Rodrigo Á, Giampiccolo D (2012) Question answering at the cross-language evaluation forum 2003-2010. Language Resources and Evaluation 46(2):177–217, DOI 10.1007/s10579-012-9177-0
2010
Earlier work this paper cites.
Ritter A, Cherry C, Dolan B (2010) Unsupervised Modeling of Twitter Conversations. In: Human Language Technologies: The 2010 Annual Conference of the North American Chapter of the Association for Computational Linguistics, Stroudsburg, PA, USA, HLT ’10, pp 172–180, URL http://dl.acm.org/citation.cfm?id=1857999.1858019
2010
Earlier work this paper cites.
Black AW, Burger S, Conkie A, Hastie H, Keizer S, Lemon O, Merigaud N, Parent G, Schubiner G, Thomson B, Williams JD, Yu K, Young S, Eskenazi M (2011) Spoken Dialog Challenge 2010: Comparison of Live and Control Test Results. In: Proceedings of the SIGDIAL 2011 Conference: The 12th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Association for Computational Linguistics, Portland, Oregon, pp 2–7
2011
Earlier work this paper cites.
Danescu C, Lee L (2011) Chameleons in Imagined Conversations: A New Approach to Understanding Coordination of Linguistic Style in Dialogs. In: Proceedings of the 2nd Workshop on Cognitive Modeling and Computational Linguistics, Association for Computational Linguistics, pp 76–87
2011
Cited alongside, same era.
DeVault D, Leuski A, Sagae K (2011) Toward Learning and Evaluation of Dialogue Policies with Text Examples. In: Proceedings of the SIGDIAL 2011 Conference: The 12th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Association for Computational Linguistics, Stroudsburg, PA, USA, pp 39–48
2011
Cited alongside, same era.
Gašić M, Jurčíček F, Thomson B, Yu K, Young S (2011) On-line policy optimisation of spoken dialogue systems via live interaction with human subjects. In: 2011 IEEE Workshop on Automatic Speech Recognition Understanding, pp 312–317, DOI 10.1109/ASRU.2011.6163950
2011
Cited alongside, same era.
Williams J, Raux A, Henderson M (2016) The Dialog State Tracking Challenge Series: A Review. Dialogue & Discourse URL https://www.microsoft.com/en-us/research/publication/the-dialog-state-tracking-challenge-series-a-review/
2016
Later among the works it cites.
Zhang X, Wang H (2016) A Joint Model of Intent Determination and Slot Filling for Spoken Language Understanding. In: Proceedings of the Twenty-Fifth International Joint Conference on Artificial Intelligence, New York, New York, USA, IJCAI’16, pp 2993–2999, URL https://www.ijcai.org/Proceedings/16/Papers/425.pdf
2016
Later among the works it cites.
Zhao T, Eskenazi M (2016) Towards End-to-End Learning for Dialog State Tracking and Management using Deep Reinforcement Learning. In: Proceedings of the SIGDIAL 2016 Conference: The 17th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Los Angeles, CA, USA, SIGDIAL’16, pp 1–10, DOI 10.18653/v1/W16-3601
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jurcícek F, Keizer S, Gasic M, Mairesse F, Thomson B, Yu K, Young SJ (2011) Real User Evaluation of Spoken Dialogue Systems Using Amazon Mechanical Turk. In: 12th Annual Conference of the International Speech Communication Association, Florence, Italy, INTERSPEECH, pp 3061–3064
2011
Cited alongside, same era.
Kolomiyets O, Moens MF (2011) A Survey on Question Answering Technology from an Information Retrieval Perspective. Information Sciences 181(24):5412–5434, DOI 10.1016/j.ins.2011.07.047
2011
Cited alongside, same era.
Lu X (2012) The relationship of lexical richness to the quality of ESL learners’ oral narratives. The Modern Language Journal 96(2):190–208, DOI 10.1111/j.1540-4781.2011.01232\_1.x
2011
Cited alongside, same era.
Ritter A, Cherry C, Dolan WB (2011) Data-driven Response Generation in Social Media. In: Proceedings of the Conference on Empirical Methods in Natural Language Processing, Edinburgh, Scotland, UK., EMNLP ’11, pp 583–593, URL http://dl.acm.org/citation.cfm?id=2145432.2145500
2011
Cited alongside, same era.
Tur G, De Mori R (2011) Spoken language understanding: Systems for extracting semantic information from speech. John Wiley & Sons
2011
Cited alongside, same era.
Tur G, Mori RD (2011) Spoken Language Understanding: Systems for Extracting Semantic Information from Speech. John Wiley & Sons
2011
Cited alongside, same era.
Zhao WX, Jiang J, Weng J, He J, Lim EP, Yan H, Li X (2011) Comparing Twitter and Traditional Media Using Topic Models. In: Proceedings of the 33rd European Conference on Advances in Information Retrieval, Springer-Verlag, Berlin, Heidelberg, ECIR’11, pp 338–349, URL http://dl.acm.org/citation.cfm?id=1996889.1996934
2011
Cited alongside, same era.
Banchs RE (2012) Movie-DiC: a Movie Dialogue Corpus for Research and Development. In: Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), Association for Computational Linguistics, pp 203–207
2012
Cited alongside, same era.
Banchs RE, Li H (2012) IRIS: a chat-oriented dialogue system based on the vector space model. In: Proceedings of the ACL 2012 Demonstrations, Jeju Island, Korea, pp 37–42
2012
Cited alongside, same era.
2017
Later among the works it cites.
Bruni E, Fernandez R (2017) Adversarial evaluation for open-domain dialogue generation. In: Proceedings of the SIGDIAL 2017 Conference: The 18th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Association for Computational Linguistics, pp 284–288
2017
Later among the works it cites.
Chen H, Liu X, Yin D, Tang J (2017) A Survey on Dialogue Systems: Recent Advances and New Frontiers. Special Interest Group on Knowledge Discovery and Data Mining (SIGKDD) Explor Newsl 19(2):25–35
2017
Later among the works it cites.
2017
Later among the works it cites.
Dubuisson Duplessis G, Charras F, Letard V, Ligozat AL, Rosset S (2017) Utterance Retrieval based on Recurrent Surface Text Patterns. In: European Conference on Information Retrieval, Aberdeen, Scotland UK, ECIR 2017, URL https://hal.archives-ouvertes.fr/hal-01436052/document
2017
Later among the works it cites.
Eric M, Krishnan L, Charette F, Manning CD (2017) Key-Value Retrieval Networks for Task-Oriented Dialogue. In: Proceedings of the SIGDIAL 2017 Conference: The 18th Annual Meeting of the Special Interest Group on Discourse and Dialogue, Saarbrücken, Germany, SIGDIAL’17, pp 37–49, DOI 10.18653/v1/W17-5506
2017
Later among the works it cites.
Hu Z, Yang Z, Liang X, Salakhutdinov R, Xing EP (2017) Toward Controlled Generation of Text. In: Proceedings of the 34th International Conference on Machine Learning, International Convention Centre, Sydney, Australia, ICML, pp 1587–1596, URL http://proceedings.mlr.press/v70/hu17e.html
2017
Later among the works it cites.
Joshi M, Choi E, Weld D, Zettlemoyer L (2017) TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Association for Computational Linguistics, Vancouver, Canada, pp 1601–1611, DOI 10.18653/v1/P17-1147
2017
Later among the works it cites.
Jurafsky D, Martin JH (2017) Speech and Language Processing, draft of 3rd edition edn, chap Dialog Systems and Chatbots
2017
Later among the works it cites.
2017
Later among the works it cites.
Mrkšić N, Ó Séaghdha D, Wen TH, Thomson B, Young S (2017) Neural Belief Tracker: Data-Driven Dialogue State Tracking. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Vancouver, Canada, ACL ’17, pp 1777–1788, DOI 10.18653/v1/P17-1163
2017
Later among the works it cites.
2017
Later among the works it cites.
Perez J, Boureau YL, Bordes A (2017) Dialog System and Technology Challenge 6 Overview of Track 1 - End-to-End Goal-Oriented Dialog learning. Tech. rep
2017
Later among the works it cites.
Sarrouti M, Ouatik El Alaoui S (2017) A Passage Retrieval Method Based on Probabilistic Information Retrieval Model and UMLS Concepts in Biomedical Question Answering. J of Biomedical Informatics 68(C):96–103, DOI 10.1016/j.jbi.2017.03.001
2017
Later among the works it cites.
Semeniuta S, Severyn A, Barth E (2017) A Hybrid Convolutional Variational Autoencoder for Text Generation. In: Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing, Copenhagen, Denmark, EMNLP, pp 627–637, URL https://www.aclweb.org/anthology/D17-1066
2017
Later among the works it cites.
Trischler A, Wang T, Yuan X, Harris J, Sordoni A, Bachman P, Suleman K (2017) NewsQA: A machine comprehension dataset. In: Proceedings of the 2nd Workshop on Representation Learning for NLP, Association for Computational Linguistics, Vancouver, Canada, pp 191–200, DOI 10.18653/v1/W17-2623
2017
Later among the works it cites.
Ultes S, Rojas Barahona LM, Su PH, Vandyke D, Kim D, Casanueva In, Budzianowski P, Mrkšić N, Wen TH, Gasic M, Young S (2017) PyDial: A Multi-domain Statistical Dialogue System Toolkit. In: Proceedings of ACL 2017, System Demonstrations, Vancouver, Canada, pp 73–78
2017
Later among the works it cites.
Wen TH, Vandyke D, Mrkšić N, Gasic M, Rojas Barahona LM, Su PH, Ultes S, Young S (2017) A Network-based End-to-End Trainable Task-oriented Dialogue System. In: Proceedings of the 15th Conference of the European Chapter of the Association for Computational Linguistics: Volume 1, Long Papers, Valencia, Spain, EACL ’17, pp 438–449, URL http://aclweb.org/anthology/E17-1042
2017
Later among the works it cites.
Xing C, Wu W, Wu Y, Liu J, Huang Y, Zhou M, Ma W (2017) Topic Aware Neural Response Generation. In: Proceedings of the Thirty-First AAAI Conference on Artificial Intelligence, San Francisco, California, USA, AAAI ’17, pp 3351–3357, URL http://aaai.org/ocs/index.php/AAAI/AAAI17/paper/view/14563
2017
Later among the works it cites.
Zhao T, Zhao R, Eskenazi M (2017) Learning discourse-level diversity for neural dialog models using conditional variational autoencoders. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Association for Computational Linguistics, Vancouver, Canada, pp 654–664, DOI 10.18653/v1/P17-1061
2017
Later among the works it cites.
Budzianowski P, Wen TH, Tseng BH, Casanueva I, Stefan U, Osman R, Gašić M (2018) MultiWOZ - A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling. In: Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP), Brussels, Belgium
2018
Later among the works it cites.
Choi E, He H, Iyyer M, Yatskar M, Yih Wt, Choi Y, Liang P, Zettlemoyer L (2018) QuAC: Question Answering in Context. In: Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP), Paris, France
2018
Later among the works it cites.
Diefenbach D, Lopez V, Singh K, Maret P (2018) Core Techniques of Question Answering Systems over Knowledge Bases: A Survey. Knowledge and Information Systems 55(3):529–569
2018
Later among the works it cites.
Furlanello T, Lipton ZC, Tschannen M, Itti L, Anandkumar A (2018) Born-again neural networks. In: Dy JG, Krause A (eds) Proceedings of the 35th International Conference on Machine Learning, ICML 2018, Stockholmsmässan, Stockholm, Sweden, July 10-15, 2018, PMLR, Proceedings of Machine Learning Research, vol 80, pp 1602–1611, URL http://proceedings.mlr.press/v80/furlanello18a.html
2018
Later among the works it cites.
Ghazvininejad M, Brockett C, Chang MW, Dolan B, Gao J, Yih Wt, Galley M (2018) A knowledge-grounded neural conversation model. In: Thirty-Second AAAI Conference on Artificial Intelligence, New Orleans, Louisiana, USA, AAAI 2018, pp 5110–5117
2018
Later among the works it cites.
Guo F, Metallinou A, Khatri C, Raju A, Venkatesh A, Ram A (2018) Topic-based evaluation for conversational bots. arXiv preprint arXiv:180103622
2018
Later among the works it cites.
Kočiský T, Schwarz J, Blunsom P, Dyer C, Hermann KM, Melis G, Grefenstette E (2018) The NarrativeQA Reading Comprehension Challenge. Transactions of the Association for Computational Linguistics 6:317–328, DOI 10.1162/tacl_a_00023
2018
Later among the works it cites.
Kreyssig F, Casanueva I, Budzianowski P, Gasic M (2018) Neural User Simulation for Corpus-based Policy Optimisation for Spoken Dialogue Systems. arXiv preprint arXiv:180506966
2018
Later among the works it cites.
Mazza R, Ambrosini L, Catenazzi N, Vanini S, Tuggener D, Tavarnesi G (2018) Behavioural Simulator For Professional Training Based On Natural Language Interaction. In: 10th International Conference on Education and New Learning Technologies, Palma, Mallorca, Spain, EDULEARN18, pp 3204–3214, URL http://repository.supsi.ch/9776/1/edulearn18-paper-lifelike.pdf
2018
Later among the works it cites.
Qu C, Yang L, Croft WB, Trippas JR, Zhang Y, Qiu M (2018) Analyzing and Characterizing User Intent in Information-seeking Conversations. In: The 41st International ACM SIGIR Conference on Research & Development in Information Retrieval, Ann Arbor, MI, USA, SIGIR 2018, pp 989–992, DOI 10.1145/3209978.3210124
2018
Later among the works it cites.
Rajpurkar P, Jia R, Liang P (2018) Know What You Don’t Know: Unanswerable Questions for SQuAD. In: Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), Association for Computational Linguistics, Melbourne, Australia, pp 784–789, DOI 10.18653/v1/P18-2124
2018
Later among the works it cites.
Reddy S, Chen D, Manning CD (2018) CoQA: A Conversational Question Answering Challenge. Transactions of the Association for Computational Linguistics 7:249–266, URL https://www.aclweb.org/anthology/Q19-1016"
2018
Later among the works it cites.
Rodrigo A, Peñas A, Miyao Y, Kando N (2018) Do systems pass university entrance exams? Information Processing & Management 54(4):564–575, DOI 10.1016/J.IPM.2018.03.002
2018
Later among the works it cites.
Saha A, Pahuja V, Khapra MM, Sankaranarayanan K, Chandar S (2018) Complex sequential question answering: Towards learning to converse over linked question answer pairs with a knowledge graph. In: McIlraith SA, Weinberger KQ (eds) Proceedings of the Thirty-Second AAAI Conference on Artificial Intelligence, (AAAI-18), the 30th innovative Applications of Artificial Intelligence (IAAI-18), and the 8th AAAI Symposium on Educational Advances in Artificial Intelligence (EAAI-18), New Orleans, Louisiana, USA, February 2-7, 2018, AAAI Press, pp 705–713, URL https://www.aaai.org/ocs/index.php/AAAI/AAAI18/paper/view/17181
2018
Later among the works it cites.
Serban IV, Lowe R, Henderson P, Charlin L, Pineau J (2018) A Survey of Available Corpora for Building Data-Driven Dialogue Systems: The Journal Version. Dialogue & Discourse 1(9)
2018
Later among the works it cites.
Talmor A, Berant J (2018) The web as a knowledge-base for answering complex questions. In: Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers), Association for Computational Linguistics, New Orleans, Louisiana, pp 641–651, DOI 10.18653/v1/N18-1059
2018
Later among the works it cites.
Tao C, Mou L, Zhao D, Yan R (2018) Ruber: An unsupervised method for automatic evaluation of open-domain dialog systems. URL https://www.aaai.org/ocs/index.php/AAAI/AAAI18/paper/view/16179/15752
2018
Later among the works it cites.
Wang A, Singh A, Michael J, Hill F, Levy O, Bowman S (2018) GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding. In: Proceedings of the 2018 {EMNLP} Workshop {B}lackbox{NLP}: Analyzing and Interpreting Neural Networks for {NLP}, Association for Computational Linguistics, Brussels, Belgium, pp 353–355, DOI 10.18653/v1/W18-5446
2018
Later among the works it cites.
Yang Y, Yih Wt, Meek C (2015) WikiQA: A challenge dataset for open-domain question answering. In: Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing, Association for Computational Linguistics, Lisbon, Portugal, pp 2013–2018, DOI 10.18653/v1/D15-1237
2018
Later among the works it cites.
Zhou L, Gao J, Li D, Shum HY (2018) The Design and Implementation of XiaoIce, an Empathetic Social Chatbot. arXiv preprint arXiv:181208989
2018
Later among the works it cites.
Byrne B, Krishnamoorthi K, Sankar C, Neelakantan A, Goodrich B, Duckworth D, Yavuz S, Dubey A, Kim K, Cedilnik A (2019) Taskmaster-1: Toward a realistic and diverse dialog dataset. In: Inui K, Jiang J, Ng V, Wan X (eds) Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing, EMNLP-IJCNLP 2019, Hong Kong, China, November 3-7, 2019, Association for Computational Linguistics, pp 4515–4524, DOI 10.18653/v1/D19-1459
2019
Closest in time.
Campos JA, Otegi A, Soroa A, Deriu J, Cieliebak M, Agirre E (2019) Conversational QA for FAQs. In: 3rd Conversational AI: “Today’s Practice and Tomorrow’s Potential” workshop at NeurIPS 2019
2019
Closest in time.
Collins E, Rozanov N, Zhang B (2019) LIDA: lightweight interactive dialogue annotator. In: Padó S, Huang R (eds) Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing, EMNLP-IJCNLP 2019, Hong Kong, China, November 3-7, 2019 - System Demonstrations, Association for Computational Linguistics, pp 121–126, DOI 10.18653/v1/D19-3021
2019
Closest in time.
Devlin J, Chang MW, Lee K, Toutanova K (2019) BERT: Pre-training of deep bidirectional transformers for language understanding. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers), Association for Computational Linguistics, Minneapolis, Minnesota, pp 4171–4186, DOI 10.18653/v1/N19-1423
2019
Closest in time.
Dušek O, Novikova J, Rieser V (2020) Evaluating the State-of-the-Art of End-to-End Natural Language Generation: The E2E NLG Challenge. Computer Speech and Language 59:123–156, DOI 10.1016/j.csl.2019.06.009
2019
Closest in time.
Gunasekara C, Kummerfeld JK, Polymenakos L, Lasecki WS (2019) DSTC7 Task 1: Noetic End-to-End Response Selection. In: 7th Edition of the Dialog System Technology Challenges at AAAI 2019, URL http://workshop.colips.org/dstc7/papers/dstc7_task1_final_report.pdf
2019
Closest in time.
Hancock B, Bordes A, Mazare PE, Weston J (2019) Learning from Dialogue after Deployment: Feed Yourself, Chatbot! In: Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, Florence, Italy, ACL 2019, pp 3667–3684, URL https://www.aclweb.org/anthology/P19-1358
2019
Closest in time.
Huang HY, Choi E, tau Yih W (2019) FlowQA: Grasping flow in history for conversational machine comprehension. In: International Conference on Learning Representations, URL https://openreview.net/forum?id=ByftGnR9KX
2019
Closest in time.
Larson S, Mahendran A, Peper JJ, Clarke C, Lee A, Hill P, Kummerfeld JK, Leach K, Laurenzano MA, Tang L, Mars J (2019) An evaluation dataset for intent classification and out-of-scope prediction. In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP), Association for Computational Linguistics, Hong Kong, China, pp 1311–1316, DOI 10.18653/v1/D19-1131
2019
Closest in time.
Lee S, Schulz H, Atkinson A, Gao J, Suleman K, Asri LE, Adada M, Huang M, Sharma S, Tay W, Li X (2019) Multi-Domain Task-Completion Dialog Challenge. Dialog System Technology Challenges 8
2019
Closest in time.
Peskov D, Clarke N, Krone J, Fodor B, Zhang Y, Youssef A, Diab M (2019) Multi-Domain Goal-Oriented Dialogues (MultiDoGO): Strategies toward Curating and Annotating Large Scale Dialogue Data. In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP), Association for Computational Linguistics, Hong Kong, China, pp 4526–4536, DOI 10.18653/v1/D19-1460
2019
Closest in time.
Qu C, Yang L, Qiu M, Zhang Y, Chen C, Croft WB, Iyyer M (2019) Attentive history selection for conversational question answering. In: Proceedings of the 28th ACM International Conference on Information and Knowledge Management, Association for Computing Machinery, New York, NY, USA, CIKM ’19, p 1391–1400, DOI 10.1145/3357384.3357905
2019
Closest in time.
Rastogi A, Zang X, Sunkara S, Gupta R, Khaitan P (2019) Towards Scalable Multi-domain Conversational Agents: The Schema-Guided Dialogue Dataset. arXiv preprint arXiv:190905855
2019
Closest in time.
Sai AB, Gupta MD, Khapra MM, Srinivasan M (2019) Re-evaluating adem: A deeper look at scoring dialogue responses. In: Proceedings of the thirty-third AAAI Conference on Artificial Intelligence, Honolulu, Hawaii, USA, AAAI’19, vol 33, pp 6220–6227, URL https://aaai.org/ojs/index.php/AAAI/article/view/4581
2019
Closest in time.
Sugiyama H, Meguro T, Higashinaka R (2019) Automatic Evaluation of Chat-Oriented Dialogue Systems Using Large-Scale Multi-references, Springer International Publishing, Cham, pp 15–25. DOI 10.1007/978-3-319-92108-2_2
2019
Closest in time.
Yang Z, Dai Z, Yang Y, Carbonell J, Salakhutdinov RR, Le QV (2019) Xlnet: Generalized autoregressive pretraining for language understanding. In: Advances in neural information processing systems, pp 5754–5764
2019
Closest in time.
Yeh YT, Chen YN (2019) FlowDelta: Modeling flow information gain in reasoning for conversational machine comprehension. In: Proceedings of the 2nd Workshop on Machine Reading for Question Answering, Association for Computational Linguistics, Hong Kong, China, pp 86–90, DOI 10.18653/v1/D19-5812
2019
Closest in time.
Adiwardana D, Luong MT, So DR, Hall J, Fiedel N, Thoppilan R, Yang Z, Kulshreshtha A, Nemade G, Lu Y, et al. (2020) Towards a human-like open-domain chatbot. arXiv preprint arXiv:200109977
2020
Closest in time.
Liu B, Tür G, Hakkani-Tür D, Shah P, Heck L (2018) Dialogue Learning with Human Teaching and Feedback in End-to-End Trainable Task-Oriented Dialogue Systems. In: Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers), Association for Computational Linguistics, New Orleans, Louisiana, USA, NAACL-HLT ’18, pp 2060–2069, URL http://aclweb.org/anthology/N18-1187
2069
Closest in time.
Galley M, Brockett C, Sordoni A, Ji Y, Auli M, Quirk C, Mitchell M, Gao J, Dolan B (2015) deltaBLEU: A Discriminative Metric for Generation Tasks with Intrinsically Diverse Targets. In: Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 2: Short Papers), Association for Computational Linguistics, ACL 2015, pp 445–450, URL http://www.aclweb.org/anthology/P15-2073
2073
Closest in time.