Fetching the paper…
Reading the bibliography…
We propose a multilingual data-driven method for generating reading comprehension questions using dependency trees.
Unified language model pre-training for natural language understanding and generation
Dong, Li, Nan Yang, Wenhui Wang, Furu Wei, Xiaodong Liu, Yu Wang, Jianfeng Gao, Ming Zhou, and Hsiao-Wuen Hon. 2019 · 1905
Earlier work this paper cites.
Cross-lingual training for automatic question generation
Kumar, Vishwajeet, Nitish Joshi, Arijit Mukherjee, Ganesh Ramakrishnan, and Preethi Jyothi. 2019 · 1906
Earlier work this paper cites.
Measuring nominal scale agreement among many raters
Fleiss, Joseph L. 1971 · 1971
Earlier work this paper cites.
The measurement of observer agreement for categorical data
Landis, J Richard and Gary G Koch. 1977 · 1977
Earlier work this paper cites.
Measures of association for cross classifications
Goodman, Leo A and William H Kruskal. 1979 · 1979
Earlier work this paper cites.
Qualitative descriptors of strength of association and effect size
Rosenthal, James A. 1996 · 1996
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, Kishore, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Analyzing quantitative data: From description to explanation
Blaikie, Norman. 2003 · 2003
Earlier work this paper cites.
Stanza: A python natural language processing toolkit for many human languages
Qi, Peng, Yuhao Zhang, Yuhui Zhang, Jason Bolton, and Christopher D Manning. 2020 · 2003
Earlier work this paper cites.
Probabilistically masked language model capable of autoregressive generation in arbitrary word order
Liao, Yi, Xin Jiang, and Qun Liu. 2020 · 2004
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, Chin-Yew. 2004 · 2004
Earlier work this paper cites.
Universal dependencies v2: An evergrowing multilingual treebank collection
Nivre, Joakim, Marie-Catherine de Marneffe, Filip Ginter, Jan Hajič, Christopher D Manning, Sampo Pyysalo, Sebastian Schuster, Francis Tyers, and Daniel Zeman. 2020 · 2004
Earlier work this paper cites.
Language models are few-shot learners
Brown, Tom B, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 2005
Earlier work this paper cites.
93 position of interrogative phrases in content questions
Dryer, Matthew S. 2005 · 2005
Earlier work this paper cites.
Free-marginal multirater kappa (multirater k [free]): An alternative to fleiss’ fixed-marginal multirater kappa
Randolph, Justus J. 2005 · 2005
Earlier work this paper cites.
Re-evaluating the role of bleu in machine translation research
Callison-Burch, Chris, Miles Osborne, and Philipp Koehn. 2006 · 2006
Earlier work this paper cites.
Meteor, m-bleu and m-ter: Evaluation metrics for high-correlation with human rankings of machine translation output
Agarwal, Abhaya and Alon Lavie. 2008 · 2008
Cited alongside, same era.
Question generation: Example of a multi-year evaluation campaign
Rus, Vasile, Zhiqiang Cai, and Art Graesser. 2008 · 2008
Cited alongside, same era.
Question generation via overgenerating transformations and ranking
Heilman, Michael and Noah A Smith. 2009 · 2009
Cited alongside, same era.
Overview of the first question generation shared task evaluation challenge
Rus, Vasile, Brendan Wyse, Paul Piwek, Mihai Lintean, Svetlana Stoyanchev, and Cristian Moldovan. 2010 · 2010
Cited alongside, same era.
Automatic question generation using discourse cues
Agarwal, Manish, Rakshit Shah, and Prashanth Mannem. 2011 · 2011
Cited alongside, same era.
How to generate cloze questions from definitions: A syntactic approach
Sharma, Shikhar, Layla El Asri, Hannes Schulz, and Jeremie Zumer. 2017 · 2017
Later among the works it cites.
Neural question generation from text: A preliminary study
Zhou, Qingyu, Nan Yang, Furu Wei, Chuanqi Tan, Hangbo Bao, and Ming Zhou. 2017 · 2017
Later among the works it cites.
Evaluation methodologies in automatic question generation 2013-2018
Amidei, Jacopo, Paul Piwek, and Alistair Willis. 2018 · 2018
Later among the works it cites.
Survey of the state of the art in natural language generation: Core tasks, applications and evaluation
Gatt, Albert and Emiel Krahmer. 2018 · 2018
Later among the works it cites.
Automatic question generation using relative pronouns and adverbs
Khullar, Payal, Konigari Rachna, Mukul Hase, and Manish Shrivastava. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gates, Donna Marie. 2011 · 2011
Cited alongside, same era.
Question generation for french: collating parsers and paraphrasing questions
Bernhard, Delphine, Louis De Viron, Véronique Moriceau, and Xavier Tannier. 2012 · 2012
Cited alongside, same era.
Generating diagnostic multiple choice comprehension cloze questions
Mostow, Jack and Hyeju Jang. 2012 · 2012
Cited alongside, same era.
Automatic generation of multiple choice questions using dependency-based semantic relations
Afzal, Naveed and Ruslan Mitkov. 2014 · 2014
Cited alongside, same era.
Meteor universal: Language specific translation evaluation for any target language
Denkowski, Michael and Alon Lavie. 2014 · 2014
Cited alongside, same era.
Leveraging multiple views of text for automatic question generation
Mazidi, Karen and Rodney D Nielsen. 2015 · 2015
Cited alongside, same era.
Cider: Consensus-based image description evaluation
Vedantam, Ramakrishna, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Later among the works it cites.
Towards a better metric for evaluating question generation systems
Nema, Preksha and Mitesh M Khapra. 2018 · 2018
Later among the works it cites.
Leveraging context information for natural question generation
Song, Linfeng, Zhiguo Wang, Wael Hamza, Yue Zhang, and Daniel Gildea. 2018 · 2018
Later among the works it cites.
Paragraph-level neural question generation with maxout pointer and gated self-attention networks
Zhao, Yao, Xiaochuan Ni, Yuanyuan Ding, and Qifa Ke. 2018 · 2018
Later among the works it cites.
Agreement is overrated: A plea for correlation to assess human evaluation reliability
Amidei, Jacopo, Paul Piwek, and Alistair Willis. 2019 · 2019
Later among the works it cites.
A recurrent bert-based model for question generation
Chan, Ying-Hong and Yao-Chung Fan. 2019 · 2019
Later among the works it cites.
Improving neural question generation using answer separation
Kim, Yanghoon, Hwanhee Lee, Joongbo Shin, and Kyomin Jung. 2019 · 2019
Later among the works it cites.
Learning to generate questions by learningwhat not to generate
Liu, Bang, Mingjun Zhao, Di Niu, Kunfeng Lai, Yancheng He, Haojie Wei, and Yu Xu. 2019 · 2019
Later among the works it cites.
Tydi qa: A benchmark for information-seeking question answering in ty pologically di verse languages
Clark, Jonathan H, Eunsol Choi, Michael Collins, Dan Garrette, Tom Kwiatkowski, Vitaly Nikolaev, and Jennimaria Palomaki. 2020 · 2020
Later among the works it cites.
Udon2: a library for manipulating universal dependencies trees
Kalpakchi, Dmytro and Johan Boye. 2020 · 2020
Later among the works it cites.
Human evaluation of automatically generated text: Current trends and best practice guidelines
van der Lee, Chris, Albert Gatt, Emiel van Miltenburg, and Emiel Krahmer. 2020 · 2020
Later among the works it cites.