Fetching the paper…
Reading the bibliography…
Research on Automatic Story Generation (ASG) relies heavily on human and automatic evaluation.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Ctrl: A conditional transformer language model for controllable generation
Nitish Shirish Keskar, Bryan McCann, Lav R Varshney, Caiming Xiong, and Richard Socher. 2019 · 1909
Earlier work this paper cites.
Better summarization evaluation with word embeddings for ROUGE
Jun-Ping Ng and Viktoria Abrecht. 2015 · 1930
Earlier work this paper cites.
A new measure of rank correlation
Maurice G Kendall. 1938 · 1938
Earlier work this paper cites.
Mathematics without numbers
John G Kemeny. 1959 · 1959
Earlier work this paper cites.
Regression Analysis
Evan J. Williams. 1959 · 1959
Earlier work this paper cites.
The measurement of observer agreement for categorical data
J Richard Landis and Gary G Koch. 1977 · 1977
Earlier work this paper cites.
Interestingness: Controlling inferences
Roger C Schank. 1978 · 1978
Earlier work this paper cites.
Tests for comparing elements of a correlation matrix
James H Steiger. 1980 · 1980
Earlier work this paper cites.
What makes a good story
Allyssa McCabe and Carole Peterson. 1984 · 1984
Earlier work this paper cites.
Story-telling as planning and learning
Michael Lebowitz. 1985 · 1985
Earlier work this paper cites.
Cognitive and affective causes of interest and liking
Asghar Iran-Nejad. 1987 · 1987
Earlier work this paper cites.
Fundamentals of social choice theory
Roger B Myerson. 1996 · 1996
Earlier work this paper cites.
Narrative intelligence and the novelty of our lives
William Lowell Randall. 1999 · 1999
Earlier work this paper cites.
Bringing stories alive: Generating interactive fiction worlds
Prithviraj Ammanabrolu, Wesley Cheung, Dan Tu, William Broniec, and Mark O. Riedl. 2020 · 2001
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
The four elements of every successful story
Robert Dickman. 2003 · 2003
Earlier work this paper cites.
Precision and recall of machine translation
I. Dan Melamed, Ryan Green, and Joseph P. Turian. 2003 · 2003
Earlier work this paper cites.
Statistical significance tests for machine translation evaluation
Philipp Koehn. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Interpreting BLEU/NIST scores: How much improvement do we need to have a better system?
Ying Zhang, Stephan Vogel, and Alex Waibel. 2004 · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Evaluating Discourse and Dialogue Coding Schemes
Richard Craggs and Mary McGee Wood. 2005 · 2005
Earlier work this paper cites.
A probabilistic interpretation of precision, recall and f-score, with implication for evaluation
Cyril Goutte and Eric Gaussier. 2005 · 2005
Earlier work this paper cites.
Evaluating evaluation methods for generation in the presence of variation
Amanda Stent, Matthew Marge, and Mohit Singhai. 2005 · 2005
Earlier work this paper cites.
Re-evaluating the role of Bleu in machine translation research
Chris Callison-Burch, Miles Osborne, and Philipp Koehn. 2006 · 2006
Earlier work this paper cites.
Narratology for interactive storytelling: A critical introduction
Marc Cavazza and David Pizzi. 2006 · 2006
Earlier work this paper cites.
Fearnot!–an emergent narrative approach to virtual dramas for anti-bullying education
Ruth Aylett, Marco Vala, Pedro Sequeira, and Ana Paiva. 2007 · 2007
Earlier work this paper cites.
Empathy and the Novel
Suzanne Keen et al. 2007 · 2007
Earlier work this paper cites.
Inter-coder agreement for computational linguistics
Ron Artstein and Massimo Poesio. 2008 · 2008
Earlier work this paper cites.
Narrative interpolation for generating and understanding stories
Su Wang, Greg Durrett, and Katrin Erk. 2020 · 2008
Earlier work this paper cites.
Coding coherence relations: Reliability and validity
Wilbert Spooren and Liesbeth Degand. 2010 · 2010
Earlier work this paper cites.
Toward supporting stories with procedurally generated game worlds
Ken Hartsook, Alexander Zook, Sauvik Das, and Mark O Riedl. 2011 · 2011
Earlier work this paper cites.
Data-driven response generation in social media
Alan Ritter, Colin Cherry, and William B. Dolan. 2011 · 2011
Earlier work this paper cites.
Computing inter-rater reliability for observational data: an overview and tutorial
Kevin A Hallgren. 2012 · 2012
Earlier work this paper cites.
Engagement via emotional heightening in" passion": On the grammatical texture of emotionally-immersive passages in short fiction
Michael Toolan. 2012 · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomás Mikolov, Ilya Sutskever, Kai Chen, Gregory S. Corrado, and Jeffrey Dean. 2013b · 2013
Cited alongside, same era.
Comparing automatic evaluation measures for image description
Desmond Elliott and Frank Keller. 2014 · 2014
Cited alongside, same era.
Testing for significance of increased correlation with human judgment
Yvette Graham and Timothy Baldwin. 2014 · 2014
Cited alongside, same era.
Borda count approximation of kemeny’s rule and pairwise voting inconsistencies
Eric Sibony. 2014 · 2014
Cited alongside, same era.
chrF: character n-gram F-score for automatic MT evaluation
Maja Popović. 2015 · 2015
Cited alongside, same era.
Cider: Consensus-based image description evaluation
Ramakrishna Vedantam, C. Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime G. Carbonell, Ruslan Salakhutdinov, and Quoc V. Le. 2019 · 2019
Later among the works it cites.
Plan-and-write: Towards better automatic storytelling
Lili Yao, Nanyun Peng, Ralph Weischedel, Kevin Knight, Dongyan Zhao, and Rui Yan. 2019 · 2019
Later among the works it cites.
MoverScore: Text generation evaluating with contextualized embeddings and earth mover distance
Wei Zhao, Maxime Peyrard, Fei Liu, Yang Gao, Christian M. Meyer, and Steffen Eger. 2019 · 2019
Later among the works it cites.
Re-evaluating evaluation in text summarization
Manik Bhandari, Pranav Narayan Gour, Atabak Ashfaq, Pengfei Liu, and Graham Neubig. 2020 · 2020
Later among the works it cites.
Modeling protagonist emotions for emotion-aware storytelling
Faeze Brahman and Snigdha Chaturvedi. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Visual storytelling
Ting-Hao Kenneth Huang, Francis Ferraro, Nasrin Mostafazadeh, Ishan Misra, Aishwarya Agrawal, Jacob Devlin, Ross Girshick, Xiaodong He, Pushmeet Kohli, Dhruv Batra, C. Lawrence Zitnick, Devi Parikh, Lucy Vanderwende, Michel Galley, and Margaret Mitchell. 2016 · 2016
Cited alongside, same era.
How NOT to evaluate your dialogue system: An empirical study of unsupervised evaluation metrics for dialogue response generation
Chia-Wei Liu, Ryan Lowe, Iulian Serban, Mike Noseworthy, Laurent Charlin, and Joelle Pineau. 2016 · 2016
Cited alongside, same era.
A corpus and cloze evaluation for deeper understanding of commonsense stories
Nasrin Mostafazadeh, Nathanael Chambers, Xiaodong He, Devi Parikh, Dhruv Batra, Lucy Vanderwende, Pushmeet Kohli, and James Allen. 2016 · 2016
Cited alongside, same era.
Imaginative, immersive and interactive engagements. the rhetoric of worldbuilding in contemporary speculative fiction
Hanna-Riikka Roine. 2016 · 2016
Cited alongside, same era.
Can machine translation systems be evaluated by the crowd alone
Yvette Graham, Timothy Baldwin, Alistair Moffat, and Justin Zobel. 2017 · 2017
Cited alongside, same era.
Argumentation mining in user-generated web discourse
Ivan Habernal and Iryna Gurevych. 2017 · 2017
Cited alongside, same era.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Later among the works it cites.
Hierarchical pre-training for sequence labelling in spoken dialog
Emile Chapuis, Pierre Colombo, Matteo Manica, Matthieu Labeau, and Chloé Clavel. 2020 · 2020
Later among the works it cites.
Guiding attention in sequence-to-sequence models for dialogue act prediction
Pierre Colombo, Emile Chapuis, Matteo Manica, Emmanuel Vignon, Giovanna Varni, and Chloé Clavel. 2020 · 2020
Later among the works it cites.
SUPERT: Towards new frontiers in unsupervised evaluation metrics for multi-document summarization
Yang Gao, Wei Zhao, and Steffen Eger. 2020 · 2020
Later among the works it cites.
Content planning for neural story generation with aristotelian rescoring
Seraphina Goldfarb-Tarrant, Tuhin Chakrabarty, Ralph Weischedel, and Nanyun Peng. 2020 · 2020
Later among the works it cites.
A knowledge-enhanced pretraining model for commonsense story generation
Jian Guan, Fei Huang, Zhihao Zhao, Xiaoyan Zhu, and Minlie Huang. 2020 · 2020
Later among the works it cites.
Heavy-tailed representations, text polarity classification & data augmentation
Hamid Jalalzai, Pierre Colombo, Chloé Clavel, Éric Gaussier, Giovanna Varni, Emmanuel Vignon, and Anne Sabourin. 2020 · 2020
Later among the works it cites.
Narrative text generation with a latent discrete plan
Harsh Jhamtani and Taylor Berg-Kirkpatrick. 2020 · 2020
Later among the works it cites.
Tangled up in BLEU: Reevaluating the evaluation of automatic machine translation evaluation metrics
Nitika Mathur, Timothy Baldwin, and Trevor Cohn. 2020 · 2020
Later among the works it cites.
PlotMachines: Outline-conditioned generation with dynamic plot state tracking
Hannah Rashkin, Asli Celikyilmaz, Yejin Choi, and Jianfeng Gao. 2020 · 2020
Later among the works it cites.
Leveraging pre-trained checkpoints for sequence generation tasks
Sascha Rothe, Shashi Narayan, and Aliaksei Severyn. 2020 · 2020
Later among the works it cites.
Fill in the BLANC: Human-free quality estimation of document summaries
Oleg Vasilyev, Vedant Dharnidharka, and John Bohannon. 2020 · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Later among the works it cites.
Bertscore: Evaluating text generation with BERT
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 2020
Later among the works it cites.
Automatic story generation: Challenges and attempts
Amal Alabdulkarim, Siyan Li, and Xiangyu Peng. 2021 · 2021
Later among the works it cites.
Automatic story generation: a survey of approaches
Arwa I Alhussain and Aqil M Azmi. 2021 · 2021
Later among the works it cites.
A preliminary survey on story interestingness: Focusing on cognitive and emotional interest
Byung-Chull Bae, Suji Jang, Youngjune Kim, and Seyoung Park. 2021 · 2021
Later among the works it cites.
Semantics of the unwritten: The effect of end of paragraph and sequence tokens on text generation with GPT2
He Bai, Peng Shi, Jimmy Lin, Luchen Tan, Kun Xiong, Wen Gao, Jie Liu, and Ming Li. 2021 · 2021
Later among the works it cites.
Learning to represent and generate text using information measures
Pierre Colombo. 2021 · 2021
Later among the works it cites.
Code-switched inspired losses for spoken dialog representations
Pierre Colombo, Emile Chapuis, Matthieu Labeau, and Chloé Clavel. 2021a · 2021
Later among the works it cites.
Beam search with bidirectional strategies for neural response generation
Pierre Colombo, Chloé Clavel, Chouchang Yack, and Giovanna Varni. 2021b · 2021
Later among the works it cites.
Automatic text evaluation through the lens of Wasserstein barycenters
Pierre Colombo, Guillaume Staerman, Chloé Clavel, and Pablo Piantanida. 2021d · 2021
Later among the works it cites.
Summeval: Re-evaluating summarization evaluation
Alexander R Fabbri, Wojciech Kryściński, Bryan McCann, Caiming Xiong, Richard Socher, and Dragomir Radev. 2021 · 2021
Later among the works it cites.
Transformer-based conditional variational autoencoder for controllable story generation
Le Fang, Tao Zeng, Chaochun Liu, Liefeng Bo, Wen Dong, and Changyou Chen. 2021 · 2021
Later among the works it cites.
The perils of using Mechanical Turk to evaluate open-ended text generation
Marzena Karpinska, Nader Akoury, and Mohit Iyyer. 2021 · 2021
Later among the works it cites.
A plug-and-play method for controlled text generation
Damian Pascual, Beni Egressy, Clara Meister, Ryan Cotterell, and Roger Wattenhofer. 2021 · 2021
Later among the works it cites.
A pseudo-metric between probability distributions based on depth-trimmed regions
Guillaume Staerman, Pavlo Mozharovskyi, Stéphan Clémençon, and Florence d’Alché Buc. 2021 · 2021
Later among the works it cites.
A temporal variational model for story generation
David Wilmot and Frank Keller. 2021 · 2021
Later among the works it cites.
Bartscore: Evaluating generated text as text generation
Weizhe Yuan, Graham Neubig, and Pengfei Liu. 2021 · 2021
Later among the works it cites.
A differential entropy estimator for training neural networks
Georg Pichler, Pierre Jean A Colombo, Malik Boudiaf, Günther Koliander, and Pablo Piantanida. 2022 · 2022
Closest in time.