Fetching the paper…
Reading the bibliography…
Automatically constructed datasets for generating text from semi-structured data (tables), such as WikiBio, often contain reference texts that diverge from the information in the corresponding semi-structured data.
Evaluating the state-of-the-art of end-to-end natural language generation: The E2E NLG Challenge
Ondřej Dušek, Jekaterina Novikova, and Verena Rieser. 2019 · 1901
Earlier work this paper cites.
Design of a knowledge-based report generator
Karen Kukich. 1983 · 1983
Earlier work this paper cites.
Text Generation: Using Discourse Strategies and Focus Constraints to Generate Natural Language Text
Kathleen R. McKeown. 1985 · 1985
Earlier work this paper cites.
Building applied natural language generation systems
Ehud Reiter and Robert Dale. 1997 · 1997
Earlier work this paper cites.
Evaluation metrics for generation
Srinivas Bangalore, Owen Rambow, and Steve Whittaker. 2000 · 2000
Earlier work this paper cites.
Automatic evaluation of machine translation quality using n-gram co-occurrence statistics
George Doddington. 2002 · 2002
Earlier work this paper cites.
Bleu: A method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Collective content selection for concept-to-text generation
Regina Barzilay and Mirella Lapata. 2005 · 2005
Earlier work this paper cites.
A probabilistic setting and lexical cooccurrence model for textual entailment
Oren Glickman and Ido Dagan. 2005 · 2005
Earlier work this paper cites.
Evaluating evaluation methods for generation in the presence of variation
Amanda Stent, Matthew Marge, and Mohit Singhai. 2005 · 2005
Earlier work this paper cites.
Comparing automatic and human evaluation of nlg systems
Anja Belz and Ehud Reiter. 2006 · 2006
Earlier work this paper cites.
Re-evaluation the role of bleu in machine translation research
Chris Callison-Burch, Miles Osborne, and Philipp Koehn. 2006 · 2006
Earlier work this paper cites.
The pascal recognising textual entailment challenge
Ido Dagan, Oren Glickman, and Bernardo Magnini. 2006 · 2006
Earlier work this paper cites.
(meta-) evaluation of machine translation
Chris Callison-Burch, Cameron Fordyce, Philipp Koehn, Christof Monz, and Josh Schroeder. 2007 · 2007
Cited alongside, same era.
Learning semantic correspondences with less supervision
Percy Liang, Michael I Jordan, and Dan Klein. 2009 · 2009
Cited alongside, same era.
PEM: A paraphrase evaluation metric exploiting parallel texts
Chang Liu, Daniel Dahlmeier, and Hwee Tou Ng. 2010 · 2010
Cited alongside, same era.
How to analyze paired comparison data
Kristi Tsukida and Maya R Gupta. 2011 · 2011
Cited alongside, same era.
Joint learning of a dual smt system for paraphrase generation
Hong Sun and Ming Zhou. 2012 · 2012
Cited alongside, same era.
A systematic comparison of smoothing techniques for sentence-level bleu
Boxing Chen and Colin Cherry. 2014 · 2014
Cited alongside, same era.
Detecting cross-lingual semantic divergence for neural machine translation
Marine Carpuat, Yogarshi Vyas, and Xing Niu. 2017 · 2017
Later among the works it cites.
Creating training corpora for micro-planners
Claire Gardent, Anastasia Shimorina, Shashi Narayan, and Laura Perez-Beltrachini. 2017 · 2017
Later among the works it cites.
Improving slot filling performance with attentive neural networks on dependency structures
Lifu Huang, Avirup Sil, Heng Ji, and Radu Florian. 2017 · 2017
Later among the works it cites.
Re-evaluating automatic metrics for image captioning
Mert Kilickaya, Aykut Erdem, Nazli Ikizler-Cinbis, and Erkut Erdem. 2017 · 2017
Later among the works it cites.
Get to the point: Summarization with pointer-generator networks
Abigail See, Peter J. Liu, and Christopher D. Manning. 2017 · 2017
Later among the works it cites.
Challenges in data-to-document generation
Sam Wiseman, Stuart M Shieber, and Alexander M Rush. 2017 · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Meteor universal: Language specific translation evaluation for any target language
Michael Denkowski and Alon Lavie. 2014 · 2014
Cited alongside, same era.
Testing for significance of increased correlation with human judgment
Yvette Graham and Timothy Baldwin. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
Cider: Consensus-based image description evaluation
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Neural text generation from structured data with application to the biography domain
Rémi Lebret, David Grangier, and Michael Auli. 2016 · 2016
Cited alongside, same era.
How not to evaluate your dialogue system: An empirical study of unsupervised evaluation metrics for dialogue response generation
Chia-Wei Liu, Ryan Lowe, Iulian Serban, Mike Noseworthy, Laurent Charlin, and Joelle Pineau. 2016 · 2016
Cited alongside, same era.
Later among the works it cites.
Optimizing statistical machine translation for text simplification
Wei Xu, Courtney Napoles, Ellie Pavlick, Quanze Chen, and Chris Callison-Burch. 2016 · 2017
Later among the works it cites.
The price of debiasing automatic metrics in natural language evalaution
Arun Chaganty, Stephen Mussmann, and Percy Liang. 2018 · 2018
Later among the works it cites.
Survey of the state of the art in natural language generation: Core tasks, applications and evaluation
Albert Gatt and Emiel Krahmer. 2018 · 2018
Later among the works it cites.
Hallucinations in neural machine translation
Katherine Lee, Orhan Firat, Ashish Agarwal, Clara Fannjiang, and David Sussillo. 2018 · 2018
Later among the works it cites.
Towards a better metric for evaluating question generation systems
Preksha Nema and Mitesh M Khapra. 2018 · 2018
Later among the works it cites.
A structured review of the validity of bleu
Ehud Reiter. 2018 · 2018
Later among the works it cites.
Object hallucination in image captioning
Anna Rohrbach, Lisa Anne Hendricks, Kaylee Burns, Trevor Darrell, and Kate Saenko. 2018 · 2018
Later among the works it cites.
Identifying semantic divergences in parallel text without annotations
Yogarshi Vyas, Xing Niu, and Marine Carpuat. 2018 · 2018
Later among the works it cites.