Fetching the paper…
Reading the bibliography…
In this paper, we study the "stability" of machine learning (ML) models within the context of larger, complex NLP systems with continuous training data updates.
A finite sample distribution-free performance bound for local discrimination rules
W. H. Rogers and T. J. Wagner. 1978 · 1978
Earlier work this paper cites.
The atis spoken language systems pilot corpus
Charles T. Hemphill, John J. Godfrey, and George R. Doddington. 1990 · 1990
Earlier work this paper cites.
Principles of risk minimization for learning theory
Vladimir Vapnik. 1991 · 1991
Earlier work this paper cites.
A simple weight decay can improve generalization
Anders Krogh and John A. Hertz. 1992 · 1992
Earlier work this paper cites.
Bagging predictors
Leo Breiman. 1996 · 1996
Earlier work this paper cites.
Algorithmic stability and sanity-check bounds for leave-one-out cross-validation
Michael J. Kearns and Dana Ron. 1999 · 1999
Earlier work this paper cites.
Ensemble methods in machine learning
Thomas G. Dietterich. 2000 · 2000
Earlier work this paper cites.
A simple algorithm for learning stable machines
Savina Andonova, André Elisseeff, Theodoros Evgeniou, and Massimiliano Pontil. 2002 · 2002
Earlier work this paper cites.
Bidirectional lstm networks for improved phoneme classification and recognition
Alex Graves, Santiago Fernández, and Jürgen Schmidhuber. 2005 · 2005
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D Manning. 2014 · 2014
Cited alongside, same era.
Comparative study between incremental and ensemble learning on data streams: Case study
Wenyu Zang, Peng Zhang, Chuan Zhou, and Li Guo. 2014 · 2014
Cited alongside, same era.
Hidden technical debt in machine learning systems
David Sculley, Gary Holt, Daniel Golovin, Eugene Davydov, Todd Phillips, Dietmar Ebner, Vinay Chaudhary, Michael Young, Jean-Francois Crespo, and Dan Dennison. 2015 · 2015
Cited alongside, same era.
Training convolutional networks with noisy labels
Sainbayar Sukhbaatar, Joan Bruna, Manohar Paluri, Lubomir Bourdev, and Rob Fergus. 2015 · 2015
Cited alongside, same era.
Launch and iterate: Reducing prediction churn
Mahdi Milani Fard, Quentin Cormier, Kevin Robert Canini, and Maya R. Gupta. 2016 · 2016
Cited alongside, same era.
All-in-1: Short text classification with one model for all languages
Barbara Plank. 2017 · 2017
Later among the works it cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Consumer complaint database check
CFPB CFPB. 2018 · 2018
Later among the works it cites.
Andrew Cotter, Heinrich Jiang, Serena Wang, Taman Narayan, Maya R. Gupta, Seungil You, and Karthik Sridharan. 2018 · 2018
Later among the works it cites.
Alice Coucke, Alaa Saade, Adrien Ball, Théodore Bluche, Alexandre Caulier, David Leroy, Clément Doumouro, Thibault Gisselbrecht, Francesco Caltagirone, Thibaut Lavril, Maël Primet, and Joseph Dureau. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tianxiang Gao and Vladimir Jojic. 2016 · 2016
Cited alongside, same era.
Satisfying real-world goals with dataset constraints
Gabriel Goh, Andrew Cotter, Maya Gupta, and Michael P Friedlander. 2016 · 2016
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals. 2016 · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Cited alongside, same era.
Making deep neural networks robust to label noise: A loss correction approach
Giorgio Patrini, Alessandro Rozza, Aditya Menon, Richard Nock, and Lizhen Qu. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Later among the works it cites.
Measuring the intrinsic dimension of objective landscapes
Chunyuan Li, Heerad Farkhoor, Rosanne Liu, and Jason Yosinski. 2018 · 2018
Later among the works it cites.
A comparison of word-based and context-based representations for classification problems in health informatics
Aditya Joshi, Sarvnaz Karimi, Ross Sparks, Cecile Paris, and C Raina MacIntyre. 2019 · 2019
Later among the works it cites.
Self-attention-based bilstm model for short text fine-grained sentiment classification
Jun Xie, Bo Chen, Xinglong Gu, Fengmei Liang, and Xinying Xu. 2019 · 2019
Later among the works it cites.
Consistency of variety of machine learning and statistical models in predicting clinical risks of individual patients: longitudinal cohort study using cardiovascular disease as exemplar
Yan Li, Matthew Sperrin, Darren M Ashcroft, and Tjeerd Pieter van Staa. 2020 · 2020
Later among the works it cites.