Fetching the paper…
Reading the bibliography…
Automatically evaluating the quality of dialogue responses for unstructured domains is a challenging problem.
Principal components analysis
Karl Pearson. 1901 · 1901
Earlier work this paper cites.
Computing machinery and intelligence
Alan M Turing. 1950 · 1950
Earlier work this paper cites.
ELIZA—a computer program for the study of natural language communication between man and machine
J. Weizenbaum. 1966 · 1966
Earlier work this paper cites.
Weighted kappa: Nominal scale agreement provision for scaled disagreement or partial credit
Jacob Cohen. 1968 · 1968
Earlier work this paper cites.
Untersuchungen zu dynamischen neuronalen netzen
Sepp Hochreiter. 1991 · 1991
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Yoshua Bengio, Patrice Simard, and Paolo Frasconi. 1994 · 1994
Earlier work this paper cites.
A new algorithm for data compression
Philip Gage. 1994 · 1994
Earlier work this paper cites.
Hierarchical recurrent neural networks for long-term dependencies
Salah El Hihi and Yoshua Bengio. 1995 · 1995
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Paradise: A framework for evaluating spoken dialogue agents
Marilyn A Walker, Diane J Litman, Candace A Kamm, and Alicia Abella. 1997 · 1997
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Memo: towards automatic usability evaluation of spoken dialogue services by user error simulations
Sebastian Möller, Roman Englert, Klaus-Peter Engelbrecht, Verena Vanessa Hafner, Anthony Jameson, Antti Oulasvirta, Alexander Raake, and Norbert Reithinger. 2006 · 2006
Earlier work this paper cites.
Regression for sentence-level mt evaluation with pseudo references
Joshua Albrecht and Rebecca Hwa. 2007 · 2007
Earlier work this paper cites.
Semi-formal evaluation of conversational characters
Ron Artstein, Sudeep Gandhe, Jillian Gerten, Anton Leuski, and David Traum. 2009 · 2009
Earlier work this paper cites.
Findings of the 2011 workshop on statistical machine translation
Chris Callison-Burch, Philipp Koehn, Christof Monz, and Omar F Zaidan. 2011 · 2011
Cited alongside, same era.
Toward learning and evaluation of dialogue policies with text examples
David DeVault, Anton Leuski, and Kenji Sagae. 2011 · 2011
Cited alongside, same era.
Data-driven response generation in social media
Alan Ritter, Colin Cherry, and William B Dolan. 2011 · 2011
Cited alongside, same era.
A systematic comparison of smoothing techniques for sentence-level bleu
Boxing Chen and Colin Cherry. 2014 · 2014
Cited alongside, same era.
Evaluating coherence in open domain conversational systems
Ryuichiro Higashinaka, Toyomi Meguro, Kenji Imamura, Hiroaki Sugiyama, Toshiro Makino, and Yoshihiro Matsuo. 2014 · 2014
Cited alongside, same era.
Neural responding machine for short-text conversation
Lifeng Shang, Zhengdong Lu, and Hang Li. 2015 · 2015
Later among the works it cites.
Results of the wmt15 metrics shared task
Miloš Stanojevic, Amir Kamran, Philipp Koehn, and Ondrej Bojar. 2015 · 2015
Later among the works it cites.
Oriol Vinyals and Quoc Le. 2015 · 2015
Later among the works it cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton. 2016 · 2016
Later among the works it cites.
Generating sentences from a continuous space
Samuel R Bowman, Luke Vilnis, Oriol Vinyals, Andrew M Dai, Rafal Jozefowicz, and Samy Bengio. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Diederik Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Results of the wmt14 metrics shared task
Matouš Machácek and Ondrej Bojar. 2014 · 2014
Cited alongside, same era.
Problematic situation analysis and automatic recognition for chi-nese online conversational system
Yang Xiang, Yaoyun Zhang, Xiaoqiang Zhou, Xiaolong Wang, and Yang Qin. 2014 · 2014
Cited alongside, same era.
deltableu: A discriminative metric for generation tasks with intrinsically diverse targets
Michel Galley, Chris Brockett, Alessandro Sordoni, Yangfeng Ji, Michael Auli, Chris Quirk, Margaret Mitchell, Jianfeng Gao, and Bill Dolan. 2015 · 2015
Cited alongside, same era.
Reval: A simple and effective machine translation evaluation metric based on recurrent neural networks
Rohit Gupta, Constantin Orasan, and Josef van Genabith. 2015 · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy. 2015 · 2015
Cited alongside, same era.
Skip-thought vectors
Ryan Kiros, Yukun Zhu, Ruslan R Salakhutdinov, Richard Zemel, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Cited alongside, same era.
Tim Cooijmans, Nicolas Ballas, César Laurent, and Aaron Courville. 2016 · 2016
Later among the works it cites.
Tweet2vec: Character-based distributed representations for social media
Bhuwan Dhingra, Zhong Zhou, Dylan Fitzpatrick, Michael Muehl, and William W Cohen. 2016 · 2016
Later among the works it cites.
Censoring representations with an adversary
Harrison Edwards and Amos Storkey. 2016 · 2016
Later among the works it cites.
A semi-automated evaluation metric for dialogue model coherence
Sudeep Gandhe and David Traum. 2016 · 2016
Later among the works it cites.
Smart reply: Automated response suggestion for email
Anjuli Kannan, Karol Kurach, Sujith Ravi, Tobias Kaufmann, Andrew Tomkins, Balint Miklos, Greg Corrado, László Lukács, Marina Ganea, Peter Young, et al. 2016 · 2016
Later among the works it cites.
Chia-Wei Liu, Ryan Lowe, Iulian V Serban, Michael Noseworthy, Laurent Charlin, and Joelle Pineau. 2016 · 2016
Later among the works it cites.
Overview of the ntcir-12 short text conversation task
Lifeng Shang, Tetsuya Sakai, Zhengdong Lu, Hang Li, Ryuichiro Higashinaka, and Yusuke Miyao. 2016 · 2016
Later among the works it cites.
Strategy and policy learning for non-task-oriented conversational systems
Zhou Yu, Ziyu Xu, Alan W Black, and Alex I Rudnicky. 2016 · 2016
Later among the works it cites.
Adversarial evaluation of dialogue models
Anjuli Kannan and Oriol Vinyals. 2017 · 2017
Closest in time.
Learning to decode for future success
Jiwei Li, Will Monroe, and Dan Jurafsky. 2017 · 2017
Closest in time.