Fetching the paper…
Reading the bibliography…
Gorman and Bedrick (2019) argued for using random splits rather than standard splits in NLP experiments.
Nonparametric estimates of standard error: the jackknife, the bootstrap and other methods
Bradley Efron. 1981 · 1981
Earlier work this paper cites.
Some statistical issues in the comparison of speech recognition algorithms
Larry Gillick and Stephen Cox. 1989 · 1989
Earlier work this paper cites.
On the computational power of neural nets
Hava Siegelmann and Eduardo Sontag. 1992 · 1992
Earlier work this paper cites.
Building a large annotated corpus of english: The penn treebank
Mitch Marcus, Beatrice Santorini, and Mary Ann Marcinkiewicz. 1993 · 1993
Earlier work this paper cites.
The lack of a priori distinctions between learning algorithms
David Wolpert. 1996 · 1996
Earlier work this paper cites.
Improving predictive inference under covariate shift by weighting the log-likelihood function
Hidetoshi Shimodaira. 2000 · 2000
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin. 2003 · 2003
Earlier work this paper cites.
Domain-specific disambiguation for typing with ambiguous keyboards
Karin Harbusch, Sasa Hasan, Hajo Hoffman, Michael Kühn, and Bernhard Schüler. 2003 · 2003
Earlier work this paper cites.
Analysis of representations for domain adaptation
Shai Ben-David, John Blitzer, Koby Crammer, and Fernando Pereira. 2006 · 2006
Earlier work this paper cites.
Integrating structured biological data by kernel maximum mean discrepancy
Karsten M. Borgwardt, Arthur Gretton, Malte J. Rasch, Hans-Peter Kriegel, Bernhard Schölkopf, and Alexander J. Smola. 2006 · 2006
Earlier work this paper cites.
Statistical comparisons of classifiers over multiple data sets
Janez Demsar. 2006 · 2006
Earlier work this paper cites.
OntoNotes: The 90% solution
Eduard Hovy, Mitchell Marcus, Martha Palmer, Lance Ramshaw, and Ralph Weischedel. 2006 · 2006
Earlier work this paper cites.
Twitter sentiment classification using distant supervision
Alec Go, Richa Bhayani, and Lei Huang. 2009 · 2009
Earlier work this paper cites.
Annotated Gigaword
Courtney Napoles, Matthew Gormley, and Benjamin Van Durme. 2012 · 2012
Earlier work this paper cites.
Overview of the 2012 shared task on parsing the web
Slav Petrov and Ryan McDonald. 2012 · 2012
Cited alongside, same era.
Estimating effect size across datasets
Anders Søgaard. 2013 · 2013
Cited alongside, same era.
Squibs: Evaluation methods for statistically dependent text
Sarvnaz Karimi, Jie Yin, and Jiri Baum. 2015 · 2015
Cited alongside, same era.
Aadam: a method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
Thang Luong, Hieu Pham, and Christopher D. Manning. 2015 · 2015
Cited alongside, same era.
A neural attention model for abstractive sentence summarization
Alexander M. Rush, Sumit Chopra, and Jason Weston. 2015 · 2015
Cited alongside, same era.
Automated scoring: Beyond natural language processing
Nitin Madnani and Aoife Cahill. 2018 · 2018
Later among the works it cites.
Evaluating neural network explanation methods using hybrid documents and morphosyntactic agreement
Nina Poerner, Benjamin Roth, and Hinrich Schütze. 2018 · 2018
Later among the works it cites.
Adversarial domain adaptation for duplicate question detection
Darsh J Shah, Tao Lei, Alessandro Moschitti, Salvatore Romeo, and Preslav Nakov. 2018 · 2018
Later among the works it cites.
Wasserstein distance guided representation learning for domain adaptation
Jian Shen, Yanru Qu, Weinan Zhang, and Yong Yu. 2018 · 2018
Later among the works it cites.
Wasserstein auto-encoders
Ilya Tolstikhin, Olivier Bousquet, Sylvain Gelly, and Bernhard Schölkopf. 2018 · 2018
Later among the works it cites.
A broad-coverage challenge corpus for sentence understanding through inference
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Abstractive text summarization using sequence-to-sequence RNNs and beyond
Ramesh Nallapati, Bowen Zhou, Cicero dos Santos, Çağlar Gu̇lçehre, and Bing Xiang. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Wasserstein gan
Martin Arjovsky, Soumith Chintala, and Leon Bottou. 2017 · 2017
Cited alongside, same era.
Six challenges for neural machine translation
Philip Koehn and Rebecca Knowles. 2017 · 2017
Cited alongside, same era.
Gec into the future: Where are we going and how do we get there?
Keisuke Sakaguchi, Courtney Napoles, and Joel Tetreault. 2017 · 2017
Cited alongside, same era.
What you can cram into a single \$&!#* vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, Germán Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Later among the works it cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Discofuse: A large-scale dataset for discourse-based sentence fusion
Mor Geva, Eric Malmi, Idan Szpektor, and Jonathan Berant. 2019 · 2019
Later among the works it cites.
We need to talk about standard splits
Kyle Gorman and Steven Bedrick. 2019 · 2019
Later among the works it cites.
Ccmatrix: Mining billions of high-quality parallel sentences on the web
Holger Schwenk, Guillaume Wenzek, Sergey Edunov, Edouard Grave, and Armand Joulin. 2019 · 2019
Later among the works it cites.
What you see is what you get: Visual pronoun coreference resolution in dialogues
Xintong Yu, Hongming Zhang, Yangqiu Song, Yan Song, and Changshui Zhang. 2019 · 2019
Later among the works it cites.
Temporally-informed analysis of named entity recognition
Shruti Rijhwani and Daniel Preotiuc-Pietro. 2020 · 2020
Closest in time.
Predictive biases in natural language processing models: A conceptual framework and overview
Deven Shah, Andrew Schwartz, and Dirk Hovy. 2020 · 2020
Closest in time.
Is the best better? Bayesian statistical model comparison for natural language processing
Piotr Szymański and Kyle Gorman. 2020 · 2020
Closest in time.