Fetching the paper…
Reading the bibliography…
Research in natural language processing proceeds, in part, by demonstrating that new models achieve superior performance (e.g., accuracy) on held-out test data, compared to previous results.
Roy Schwartz, Jesse Dodge, Noah A. Smith, and Oren Etzioni. 2019 · 1907
Earlier work this paper cites.
An Introduction to the Bootstrap
Bradley Efron and Robert Tibshirani. 1994 · 1994
Earlier work this paper cites.
Why most published research findings are false
John P. A. Ioannidis. 2005 · 2005
Earlier work this paper cites.
Neutralizing linguistically problematic annotations in unsupervised dependency parsing evaluation
Roy Schwartz, Omri Abend, Roi Reichart, and Ari Rappoport. 2011 · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio. 2012 · 2012
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D. Manning, Andrew Y. Ng, and Christopher Potts. 2013 · 2013
Earlier work this paper cites.
The statistical crisis in science
Andrew Gelman and Eric Loken. 2014 · 2014
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Bayesian optimization of text representations
Dani Yogatama and Noah A. Smith. 2015 · 2015
Earlier work this paper cites.
A decomposable attention model for natural language inference
Ankur P. Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit. 2016 · 2016
Earlier work this paper cites.
SQuAD: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Earlier work this paper cites.
Enhanced LSTM for natural language inference
Qian Chen, Xiao-Dan Zhu, Zhen-Hua Ling, Si Wei, Hui Jiang, and Diana Inkpen. 2017 · 2017
Earlier work this paper cites.
Replicability analysis for natural language processing: Testing significance with multiple datasets
Rotem Dror, Gili Baumer, Marina Bogomolov, and Roi Reichart. 2017 · 2017
Earlier work this paper cites.
Hyperband: Bandit-based configuration evaluation for hyperparameter optimization
Lisha Li, Kevin Jamieson, Giulia DeSalvo, Afshin Rostamizadeh, and Ameet Talwalkar. 2017 · 2017
Cited alongside, same era.
Learned in translation: Contextualized word vectors
Bryan McCann, James Bradbury, Caiming Xiong, and Richard Socher. 2017 · 2017
Cited alongside, same era.
Nils Reimers and Iryna Gurevych. 2017 · 2017
Cited alongside, same era.
Bidirectional attention flow for machine comprehension
Min Joon Seo, Aniruddha Kembhavi, Ali Farhadi, and Hannaneh Hajishirzi. 2017 · 2017
Cited alongside, same era.
AllenNLP: A deep semantic natural language processing platform
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke S. Zettlemoyer. 2018 · 2018
Later among the works it cites.
Sentence encoders on STILTs: Supplementary training on intermediate labeled-data tasks
Jason Phang, Thibault Févry, and Samuel R. Bowman. 2018 · 2018
Later among the works it cites.
Winner’s curse? On pace, progress, and empirical rigor
D. Sculley, Jasper Snoek, Ali Rahimi, and Alex Wiltschko. 2018 · 2018
Later among the works it cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R. Bowman. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Matt Gardner, Joel Grus, Mark Neumann, Oyvind Tafjord, Pradeep Dasigi, Nelson F. Liu, Matthew E. Peters, Michael Schmitz, and Luke S. Zettlemoyer. 2018 · 2018
Cited alongside, same era.
Timnit Gebru, Jamie H. Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna M. Wallach, Hal Daumé, and Kate Crawford. 2018 · 2018
Cited alongside, same era.
State of the art: Reproducibility in artificial intelligence
Odd Erik Gundersen and Sigbjørn Kjensmo. 2018 · 2018
Cited alongside, same era.
Deep reinforcement learning that matters
Peter Henderson, Riashat Islam, Philip Bachman, Joelle Pineau, Doina Precup, and David Meger. 2018 · 2018
Cited alongside, same era.
SciTaiL: A textual entailment dataset from science question answering
Tushar Khot, Ashutosh Sabharwal, and Peter Clark. 2018 · 2018
Cited alongside, same era.
Tune: A research platform for distributed model selection and training
Richard Liaw, Eric Liang, Robert Nishihara, Philipp Moritz, Joseph E Gonzalez, and Ion Stoica. 2018 · 2018
Cited alongside, same era.
Troubling trends in machine learning scholarship
Zachary C. Lipton and Jacob Steinhardt. 2018 · 2018
Cited alongside, same era.
Are GANs created equal? A large-scale study
Mario Lucic, Karol Kurach, Marcin Michalski, Olivier Bousquet, and Sylvain Gelly. 2018 · 2018
Cited alongside, same era.
Ye Zhang and Byron Wallace. 2015 · 2018
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Closest in time.
We need to talk about standard splits
Kyle Gorman and Steven Bedrick. 2019 · 2019
Closest in time.
Random search and reproducibility for neural architecture search
Liam Li and Ameet Talwalkar. 2019 · 2019
Closest in time.
Model cards for model reporting
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru. 2019 · 2019
Closest in time.
To tune or not to tune? Adapting pretrained representations to diverse tasks
Matthew Peters, Sebastian Ruder, and Noah A. Smith. 2019 · 2019
Closest in time.
Machine learning reproducibility checklist
Joelle Pineau. 2019 · 2019
Closest in time.
How the transformers broke NLP leaderboards
Anna Rogers. 2019 · 2019
Closest in time.
Energy and policy considerations for deep learning in NLP
Emma Strubell, Ananya Ganesh, and Andrew McCallum. 2019 · 2019
Closest in time.