Fetching the paper…
Reading the bibliography…
We reassess a recent study (Hassan et al., 2018) that claimed that machine translation (MT) has reached human parity for the translation of news from Chinese into English, using pairwise ranking and considering three variables that were not taken into account in that previous study: the language in which the source side of the test set was originally written, the translation proficiency of the evaluators, and the provision of inter-sentential context.
A Coefficient of Agreement for Nominal Scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
Universals of translation
Sara Laviosa-Braithwaite. 1998 · 1998
Earlier work this paper cites.
Fast, cheap, and creative: Evaluating translation quality using Amazon’s Mechanical Turk
Chris Callison-Burch. 2009 · 2009
Earlier work this paper cites.
Appraise: An open-source toolkit for manual evaluation of machine translation output
Christian Federmann. 2012 · 2012
Earlier work this paper cites.
Towards a literary machine translation: The role of referential cohesion
Rob Voigt and Dan Jurafsky. 2012 · 2012
Earlier work this paper cites.
Findings of the 2013 Workshop on Statistical Machine Translation
Ondřej Bojar, Christian Buck, Chris Callison-Burch, Christian Federmann, Barry Haddow, Philipp Koehn, Christof Monz, Matt Post, Radu Soricut, and Lucia Specia. 2013 · 2013
Earlier work this paper cites.
Efficient Elicitation of Annotations for Human Evaluation of Machine Translation
Keisuke Sakaguchi, Matt Post, and Benjamin Van Durme. 2014 · 2014
Earlier work this paper cites.
Accurate evaluation of segment-level machine translation metrics
Yvette Graham, Timothy Baldwin, and Nitika Mathur. 2015 · 2015
Cited alongside, same era.
Neural versus phrase-based machine translation quality: a case study
Luisa Bentivogli, Arianna Bisazza, Mauro Cettolo, and Marcello Federico. 2016 · 2016
Cited alongside, same era.
Findings of the 2016 conference on machine translation
Ondřej Bojar, Rajen Chatterjee, Christian Federmann, Yvette Graham, Barry Haddow, Matthias Huck, Antonio Jimeno Yepes, Philipp Koehn, Varvara Logacheva, Christof Monz, Matteo Negri, Aurelie Neveol, Mariana Neves, Martin Popel, Matt Post, Raphael Rubino, Carolina Scarton, Lucia Specia, Marco Turchi, Karin Verspoor, and Marcos Zampieri. 2016 · 2016
Cited alongside, same era.
Edinburgh Neural Machine Translation Systems for WMT 16
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Lukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Greg Corrado, Macduff Hughes, and Jeffrey Dean. 2016 · 2016
A multifaceted evaluation of neural versus phrase-based machine translation for 9 language directions
Antonio Toral and Víctor M. Sánchez-Cartagena. 2017 · 2017
Later among the works it cites.
Exploiting cross-sentence context for neural machine translation
Longyue Wang, Zhaopeng Tu, Andy Way, and Liu Qun. 2017a · 2017
Later among the works it cites.
Sogou neural machine translation systems for WMT17
Yuguang Wang, Shanbo Cheng, Liyang Jiang, Jiajun Yang, Wei Chen, Muze Li, Lin Shi, Yanfeng Wang, and Hongtao Yang. 2017b · 2017
Later among the works it cites.
Achieving Human Parity on Automatic Chinese to English News Translation
Hany Hassan, Anthony Aue, Chang Chen, Vishal Chowdhary, Jonathan Clark, Christian Federmann, Xuedong Huang, Marcin Junczys-Dowmunt, Will Lewis, Mu Li, Shujie Liu, Tie-Yan Liu, Renqian Luo, Arul Menezes, Tao Qin, Frank Seide, Xu Tan, Fei Tian, Lijun Wu, Shuangzhi Wu, Yingce Xia, Dongdong Zhang, Zhirui Zhang, and Ming Zhou. 2018 · 2018
Closest in time.
Has Machine Translation Achieved Human Parity? A Case for Document-level Evaluation
Samuel Läubli, Rico Sennrich, and Martin Volk. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Findings of the 2017 conference on machine translation (wmt17)
Ondřej Bojar, Rajen Chatterjee, Christian Federmann, Yvette Graham, Barry Haddow, Shujian Huang, Matthias Huck, Philipp Koehn, Qun Liu, Varvara Logacheva, Christof Monz, Matteo Negri, Matt Post, Raphael Rubino, Lucia Specia, and Marco Turchi. 2017 · 2017
Cited alongside, same era.
A Comparative Quality Evaluation of PBSMT and NMT using Professional Translators
Sheila Castilho, Joss Moorkens, Federico Gaspari, Rico Sennrich, Vilelmini Sosoni, Panayota Georgakopoulou, Pintu Lohar, Andy Way, Antonio Valerio Miceli Barone, and Maria Gialama. 2017b · 2017
Cited alongside, same era.
Is neural machine translation the new state of the art?
Sheila Castilho, Joss Moorkens, Federico Gaspari, Iacer Calixto, John Tinsley, and Andy Way. 2017a
Cited in the paper.
Crowdsourcing for NMT evaluation: Professional translators versus the crowd
Sheila Castilho, Joss Moorkens, Federico Gaspari, Andy Way, Panayota Georgakopoulou, Maria Gialama, Vilelmini Sosoni, and Rico Sennrich. 2017c
Cited in the paper.
Closest in time.
Machine translation: Where are we at today?
Andy Way. 2018 · 2018
Closest in time.