Fetching the paper…
Reading the bibliography…
Open Domain dialog system evaluation is one of the most important challenges in dialog research.
Beyond turing: Intelligent agents centered on the user
Maxine Eskénazi, Shikib Mehri, Evgeniia Razumovskaia, and Tiancheng Zhao. 2019 · 1901
Earlier work this paper cites.
Sanghyun Yi, Rahul Goel, Chandra Khatri, Tagyoung Chung, Behnam Hedayatnia, Anu Venkatesh, Raefer Gabriel, and Dilek Hakkani-Tür. 2019 · 1904
Earlier work this paper cites.
When does label smoothing help?
Rafael Müller, Simon Kornblith, and Geoffrey E. Hinton. 2019 · 1906
Earlier work this paper cites.
Investigating evaluation of open-domain dialogue systems with human generated multiple references
Prakhar Gupta, Shikib Mehri, Tiancheng Zhao, Amy Pavel, Maxine Eskénazi, and Jeffrey P. Bigham. 2019 · 1907
Earlier work this paper cites.
ACUTE-EVAL: improved dialogue evaluation with optimized questions and multi-turn comparisons
Margaret Li, Jason Weston, and Stephen Roller. 2019a · 1909
Earlier work this paper cites.
MOSS: end-to-end dialog system framework with modular supervision
Weixin Liang, Youzhi Tian, Chengcai Chen, and Zhou Yu. 2019 · 1909
Earlier work this paper cites.
Litgen: Genetic literature recommendation guided by human explanations
Allen Nie, Arturo L. Pineda, Matt W. Wright Hannah Wand, Bryan Wulf, Helio A. Costa, Ronak Y. Patel, Carlos D. Bustamante, and James Zou. 2019 · 1909
Earlier work this paper cites.
Who’s responsible? jointly quantifying the contribution of the learning algorithm and training data
Gal Yona, Amirata Ghorbani, and James Zou. 2019 · 1910
Earlier work this paper cites.
End-to-end trainable non-collaborative dialog system
Yu Li, Kun Qian, Weiyan Shi, and Zhou Yu. 2019b · 1911
Earlier work this paper cites.
Advcodec: Towards A unified framework for adversarial text generation
Boxin Wang, Hengzhi Pei, Han Liu, and Bo Li. 2019a · 1912
Earlier work this paper cites.
Weighted kappa: Nominal scale agreement provision for scaled disagreement or partial credit
Jacob Cohen. 1968 · 1968
Earlier work this paper cites.
On the uniqueness of the shapley value
Pradeep Dubey. 1975 · 1975
Earlier work this paper cites.
Bargaining foundations of shapley value
Faruk Gul. 1989 · 1989
Earlier work this paper cites.
Measuring customer satisfaction: fact and artifact
Robert A Peterson and William R Wilson. 1992 · 1992
Earlier work this paper cites.
A comparison of question scales used for measuring customer satisfaction
Peter J Danaher and Vanessa Haddrell. 1996 · 1996
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Semi-supervised models via data augmentationfor classifying interactive affective responses
Jiaao Chen, Yuwei Wu, and Diyi Yang. 2020 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
METEOR: an automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Learning to rank using gradient descent
Christopher J. C. Burges, Tal Shaked, Erin Renshaw, Ari Lazier, Matt Deeds, Nicole Hamilton, and Gregory N. Hullender. 2005 · 2005
Earlier work this paper cites.
Feature selection based on the shapley value
Shay B Cohen, Eytan Ruppin, and Gideon Dror. 2005 · 2005
Earlier work this paper cites.
Graph-based, self-supervised program repair from diagnostic feedback
Michihiro Yasunaga and Percy Liang. 2020 · 2005
Cited alongside, same era.
Learning to rank with nonsmooth cost functions
Christopher J. C. Burges, Robert Ragno, and Quoc Viet Le. 2006 · 2006
Cited alongside, same era.
The relative incidence of positive and negative word of mouth: A multi-category study
Robert East, Kathy Hammond, and Malcolm Wright. 2007 · 2007
Cited alongside, same era.
Learning to predict engagement with a spoken dialog system in open-world settings
Dan Bohus and Eric Horvitz. 2009 · 2009
Cited alongside, same era.
Preference learning
Johannes Fürnkranz and Eyke Hüllermeier. 2010 · 2010
Cited alongside, same era.
Sequential and temporal dynamics of online opinion
David Godes and José C Silva. 2012 · 2012
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Quantifying facial age by posterior of age comparisons
Yunxuan Zhang, Li Liu, Cheng Li, and Chen Change Loy. 2017 · 2017
Later among the works it cites.
Gunrock: Building a human-like social bot by leveraging large scale real user data
Chun-Yen Chen, Dian Yu, Weiming Wen, Yi Mang Yang, Jiaping Zhang, Mingyang Zhou, Kevin Jesse, Austin Chau, Antara Bhowmick, Shreenath Iyer, et al. 2018 · 2018
Later among the works it cites.
Sounding board: A user-centric and content-driven social chatbot
Hao Fang, Hao Cheng, Maarten Sap, Elizabeth Clark, Ari Holtzman, Yejin Choi, Noah A. Smith, and Mari Ostendorf. 2018 · 2018
Later among the works it cites.
Topic-based evaluation for conversational bots
Fenfei Guo, Angeliki Metallinou, Chandra Khatri, Anirudh Raju, Anu Venkatesh, and Ashwin Ram. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A comparison of greedy and optimal assessment of natural language student input using word-to-word similarity metrics
Vasile Rus and Mihai C. Lintean. 2012 · 2012
Cited alongside, same era.
Bootstrapping dialog systems with word embeddings
Gabriel Forgues, Joelle Pineau, Jean-Marie Larchevêque, and Réal Tremblay. 2014 · 2014
Cited alongside, same era.
Similarity comparisons for interactive fine-grained categorization
Catherine Wah, Grant Van Horn, Steve Branson, Subhransu Maji, Pietro Perona, and Serge Belongie. 2014 · 2014
Cited alongside, same era.
DEX: deep expectation of apparent age from a single image
Rasmus Rothe, Radu Timofte, and Luc Van Gool. 2015 · 2015
Cited alongside, same era.
Oriol Vinyals and Quoc V. Le. 2015 · 2015
Cited alongside, same era.
A first look at online reputation on airbnb, where every stay is above average
Georgios Zervas, Davide Proserpio, and John Byers. 2015 · 2015
Cited alongside, same era.
Importance of a search strategy in neural dialogue modelling
Ilya Kulikov, Alexander H. Miller, Kyunghyun Cho, and Jason Weston. 2018 · 2018
Later among the works it cites.
Memcloak: Practical access obfuscation for untrusted memory
Weixin Liang, Kai Bu, Ke Li, Jinhong Li, and Arya Tavakoli. 2018 · 2018
Later among the works it cites.
The extreme distribution of online reviews: Prevalence, drivers and implications
Verena Schoenmüller, Oded Netzer, and Florian Stahl. 2018 · 2018
Later among the works it cites.
Finding convincing arguments using scalable bayesian preference learning
Edwin D. Simpson and Iryna Gurevych. 2018 · 2018
Later among the works it cites.
RUBER: an unsupervised method for automatic evaluation of open-domain dialog systems
Chongyang Tao, Lili Mou, Dongyan Zhao, and Rui Yan. 2018 · 2018
Later among the works it cites.
On evaluating and comparing conversational agents
Anu Venkatesh, Chandra Khatri, Ashwin Ram, Fenfei Guo, Raefer Gabriel, Ashish Nagar, Rohit Prasad, Ming Cheng, Behnam Hedayatnia, Angeliki Metallinou, Rahul Goel, Shaohua Yang, and Anirudh Raju. 2018 · 2018
Later among the works it cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Cu-net: Component unmixing network for textile fiber identification
Zunlei Feng, Weixin Liang, Daocheng Tao, Li Sun, Anxiang Zeng, and Mingli Song. 2019 · 2019
Later among the works it cites.
Data shapley: Equitable valuation of data for machine learning
Amirata Ghorbani and James Y. Zou. 2019 · 2019
Later among the works it cites.
Music transformer: Generating music with long-term structure
Cheng-Zhi Anna Huang, Ashish Vaswani, Jakob Uszkoreit, Ian Simon, Curtis Hawthorne, Noam Shazeer, Andrew M. Dai, Matthew D. Hoffman, Monica Dinculescu, and Douglas Eck. 2019 · 2019
Later among the works it cites.
Deepstore: In-storage acceleration for intelligent queries
Vikram Sharma Mailthody, Zaid Qureshi, Weixin Liang, Ziyan Feng, Simon Garcia De Gonzalo, Youjie Li, Hubertus Franke, Jinjun Xiong, Jian Huang, and Wen-Mei Hwu. 2019 · 2019
Later among the works it cites.
What makes a good counselor? learning to distinguish between high-quality and low-quality counseling conversations
Verónica Pérez-Rosas, Xinyi Wu, Kenneth Resnicow, and Rada Mihalcea. 2019 · 2019
Later among the works it cites.
Gunrock: A social bot for complex and engaging long conversations
Dian Yu, Michelle Cohn, Yi Mang Yang, Chun-Yen Chen, Weiming Wen, Jiaping Zhang, Mingyang Zhou, Kevin Jesse, Austin Chau, Antara Bhowmick, Shreenath Iyer, Giritheja Sreenivasulu, Sam Davidson, Ashwin Bhandare, and Zhou Yu. 2019 · 2019
Later among the works it cites.
Xuandong Zhao, Xiang Li, Ning Guo, Zhiling Zhou, Xiaxia Meng, and Quanzheng Li. 2019 · 2019
Later among the works it cites.
Bond: Bert-assisted open-domain named entity recognition with distant supervision
Chen Liang, Yue Yu, Haoming Jiang, Siawpeng Er, Ruijia Wang, Tuo Zhao, and Chao Zhang. 2020 · 2020
Closest in time.
Steam: Self-supervised taxonomy expansion with mini-paths
Yue Yu, Yinghao Li, Jiaming Shen, Hao Feng, Jimeng Sun, and Chao Zhang. 2020 · 2020
Closest in time.