Fetching the paper…
Reading the bibliography…
Metric validation in Grammatical Error Correction (GEC) is currently done by observing the correlation between human and metric-induced rankings.
Time Warps, String Edits, and Macromolecules: The Theory and Practice of Sequence Comparison
Joseph B Kruskal and David Sankoff. 1983 · 1983
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Developing an open-source, rule-based proofreading tool
Marcin Miłkowski. 2010 · 2010
Earlier work this paper cites.
A grain of salt for the wmt manual evaluation
Ondřej Bojar, Miloš Ercegovčević, Martin Popel, and Omar F Zaidan. 2011 · 2011
Earlier work this paper cites.
Better evaluation for grammatical error correction
Daniel Dahlmeier and Hwee Tou Ng. 2012 · 2012
Earlier work this paper cites.
Hoo 2012: A report on the preposition and determiner error correction shared task
Robert Dale, Ilya Anisimoff, and George Narroway. 2012 · 2012
Earlier work this paper cites.
Measurement of progress in machine translation
Yvette Graham, Timothy Baldwin, Aaron Harwood, Alistair Moffat, and Justin Zobel. 2012 · 2012
Earlier work this paper cites.
Simulating human judgment in machine translation evaluation campaigns
Philipp Koehn. 2012 · 2012
Earlier work this paper cites.
Putting human assessments of machine translation systems in order
Adam Lopez. 2012 · 2012
Earlier work this paper cites.
Joint learning of a dual smt system for paraphrase generation
Hong Sun and Ming Zhou. 2012 · 2012
Earlier work this paper cites.
Building a large annotated corpus of learner english: The nus corpus of learner english
Daniel Dahlmeier, Hwee Tou Ng, and Siew Mei Wu. 2013 · 2013
Earlier work this paper cites.
Findings of the 2014 workshop on statistical machine translation
Ondrej Bojar, Christian Buck, Christian Federmann, Barry Haddow, Philipp Koehn, Johannes Leveling, Christof Monz, Pavel Pecina, Matt Post, Herve Saint-Amand, et al. 2014 · 2014
Earlier work this paper cites.
A systematic comparison of smoothing techniques for sentence-level bleu
Boxing Chen and Colin Cherry. 2014 · 2014
Cited alongside, same era.
The conll-2014 shared task on grammatical error correction
Hwee Tou Ng, Siew Mei Wu, Ted Briscoe, Christian Hadiwinoto, Raymond Hendy Susanto, and Christopher Bryant. 2014 · 2014
Cited alongside, same era.
Efficient elicitation of annotations for human evaluation of machine translation
Keisuke Sakaguchi, Matt Post, and Benjamin Van Durme. 2014 · 2014
Cited alongside, same era.
How far are we from fully automatic high quality grammatical error correction?
Christopher Bryant and Hwee Tou Ng. 2015 · 2015
Cited alongside, same era.
Evaluating human pairwise preference judgments
Mark Dras. 2015 · 2015
Cited alongside, same era.
Towards a standard evaluation method for grammatical error detection and correction
Mariano Felice and Ted Briscoe. 2015 · 2015
Automatic extraction of learner errors in esl sentences using linguistically enhanced alignments
Mariano Felice, Christopher Bryant, and Ted Briscoe. 2016 · 2016
Later among the works it cites.
There’s no comparison: Reference-less evaluation metrics in grammatical error correction
Courtney Napoles, Keisuke Sakaguchi, and Joel Tetreault. 2016b · 2016
Later among the works it cites.
Reassessing the goals of grammatical error correction: Fluency instead of grammaticality
Keisuke Sakaguchi, Courtney Napoles, Matt Post, and Joel Tetreault. 2016 · 2016
Later among the works it cites.
Optimizing statistical machine translation for text simplification
Wei Xu, Courtney Napoles, Ellie Pavlick, Quanze Chen, and Chris Callison-Burch. 2016 · 2016
Later among the works it cites.
Reference-based metrics can be replaced with reference-less metrics in evaluating grammatical error correction systems
Hiroki Asano, Tomoya Mizumoto, and Kentaro Inui. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Accurate evaluation of segment-level machine translation metrics
Yvette Graham, Timothy Baldwin, and Nitika Mathur. 2015 · 2015
Cited alongside, same era.
Human evaluation of grammatical error correction systems
Roman Grundkiewicz, Marcin Junczys-Dowmunt, Edward Gillian, et al. 2015 · 2015
Cited alongside, same era.
Ground truth for grammatical error correction metrics
Courtney Napoles, Keisuke Sakaguchi, Matt Post, and Joel Tetreault. 2015 · 2015
Cited alongside, same era.
Hume: Human ucca-based evaluation of machine translation
Alexandra Birch, Omri Abend, Ondřej Bojar, and Barry Haddow. 2016 · 2016
Cited alongside, same era.
Results of the wmt16 metrics shared task
Ondřej Bojar, Yvette Graham, Amir Kamran, and Miloš Stanojević. 2016 · 2016
Cited alongside, same era.
Inherent biases in reference-based evaluation for grammatical error correction and text simplification
Leshem Choshen and Omri Abend. 2018a
Cited in the paper.
Findings of the 2017 conference on machine translation (wmt17)
Ondřej Bojar, Rajen Chatterjee, Christian Federmann, Yvette Graham, Barry Haddow, Shujian Huang, Matthias Huck, Philipp Koehn, Qun Liu, Varvara Logacheva, et al. 2017 · 2017
Later among the works it cites.
Automatic annotation and evaluation of error types for grammatical error correction
Christopher Bryant, Mariano Felice, and Ted Briscoe. 2017 · 2017
Later among the works it cites.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
Ranjay Krishna, Yuke Zhu, Oliver Groth, Justin Johnson, Kenji Hata, Joshua Kravitz, Stephanie Chen, Yannis Kalantidis, Li-Jia Li, David A Shamma, et al. 2017 · 2017
Later among the works it cites.
Nematus: a toolkit for neural machine translation
Rico Sennrich, Orhan Firat, Kyunghyun Cho, Alexandra Birch, Barry Haddow, Julian Hitschler, Marcin Junczys-Dowmunt, Samuel Läubli, Antonio Valerio Miceli Barone, Jozef Mokry, et al. 2017 · 2017
Later among the works it cites.
Seqgan: Sequence generative adversarial nets with policy gradient
Lantao Yu, Weinan Zhang, Jun Wang, and Yong Yu. 2017 · 2017
Later among the works it cites.
Reference-less measure of faithfulness for grammatical error correction
Leshem Choshen and Omri Abend. 2018b · 2018
Closest in time.