Fetching the paper…
Reading the bibliography…
We present SacreROUGE, an open-source library for using and developing summarization evaluation metrics.
Bertscore: Evaluating text generation with BERT
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2019 · 1904
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Summarization system evaluation revisited: N-gram graphs
George Giannakopoulos, Vangelis Karkaletsis, George A. Vouros, and Panagiotis Stamatopoulos. 2008 · 2008
Earlier work this paper cites.
Summarization evaluation using transformed basic elements
Stephen Tratz and Eduard H. Hovy. 2008 · 2008
Earlier work this paper cites.
Automatically evaluating content selection in summarization without human models
Annie Louis and Ani Nenkova. 2009 · 2009
Earlier work this paper cites.
Summarization system evaluation variations based on n-gram graphs
George Giannakopoulos and Vangelis Karkaletsis. 2010 · 2010
Cited alongside, same era.
Summary evaluation: Together we stand npower-ed
George Giannakopoulos and Vangelis Karkaletsis. 2013 · 2013
Cited alongside, same era.
Meteor universal: Language specific translation evaluation for any target language
Michael J. Denkowski and Alon Lavie. 2014 · 2014
Cited alongside, same era.
Opennmt: Open-source toolkit for neural machine translation
Guillaume Klein, Yoon Kim, Yuntian Deng, Jean Senellart, and Alexander M. Rush. 2017 · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Cited alongside, same era.
The price of debiasing automatic metrics in natural language evalaution
Arun Chaganty, Stephen Mussmann, and Percy Liang. 2018 · 2018
Later among the works it cites.
AllenNLP: A deep semantic natural language processing platform
Matt Gardner, Joel Grus, Mark Neumann, Oyvind Tafjord, Pradeep Dasigi, Nelson F. Liu, Matthew Peters, Michael Schmitz, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
A call for clarity in reporting BLEU scores
Matt Post. 2018 · 2018
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
SUM-QE: a bert-based summary quality estimation model
Stratos Xenouleas, Prodromos Malakasiotis, Marianna Apidianaki, and Ion Androutsopoulos. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…