Fetching the paper…
Reading the bibliography…
Deployed machine learning models should be updated to take advantage of a larger sample size to improve performance, as more data is gathered over time.
Stacked generalization
David H Wolpert · 1992
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y Lecun, L Bottou, Y Bengio, and P Haffner · 1998
Earlier work this paper cites.
MNIST handwritten digit database
Yann LeCun and Corinna Cortes · 2010
Earlier work this paper cites.
An analysis of Single-Layer networks in unsupervised feature learning
Adam Coates, Andrew Ng, and Honglak Lee · 2011
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Andrew L Maas, Raymond E Daly, Peter T Pham, Dan Huang, Andrew Y Ng, and Christopher Potts · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Yuval Netzer, Tao Wang, Adam Coates, Alessandro Bissacco, Bo Wu, and Andrew Y Ng · 2011
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Earlier work this paper cites.
A review on evaluation metrics for data classification evaluations
Mohammad Hossin and M.N Sulaiman · 2015
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun · 2015
Earlier work this paper cites.
Launch and iterate: Reducing prediction churn
Mahdi Milani Fard, Quentin Cormier, Kevin Canini, and Maya Gupta · 2016
Earlier work this paper cites.
Equality of opportunity in supervised learning
Moritz Hardt, Eric Price, and Nathan Srebro · 2016
Earlier work this paper cites.
Emnist: Extending mnist to handwritten letters
Gregory Cohen, Saeed Afshar, Jonathan Tapson, and André van Schaik · 2017
Earlier work this paper cites.
UCI machine learning repository, 2017
Dheeru Dua and Casey Graff · 2017
Earlier work this paper cites.
On calibration of modern neural networks
Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q Weinberger · 2017
Cited alongside, same era.
Exploring generalization in deep learning
Behnam Neyshabur, Srinadh Bhojanapalli, David McAllester, and Nathan Srebro · 2017
Cited alongside, same era.
Fashion-MNIST: a novel image dataset for benchmarking machine learning algorithms
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Cited alongside, same era.
Large scale distributed neural network training through online distillation
Rohan Anil, Gabriel Pereyra, Alexandre Passos, Robert Ormandi, George E Dahl, and Geoffrey E Hinton · 2018
Cited alongside, same era.
Deep learning for classical japanese literature
Tarin Clanuwat, Mikel Bober-Irizar, Asanobu Kitamoto, Alex Lamb, Kazuaki Yamamoto, and David Ha · 2018
Cited alongside, same era.
The surprising simplicity of the early-time learning dynamics of neural networks
Wei Hu, Lechao Xiao, Ben Adlam, and Jeffrey Pennington · 2020
Later among the works it cites.
Energy-based out-of-distribution detection
Weitang Liu, Xiaoyun Wang, John D Owens, and Yixuan Li · 2020
Later among the works it cites.
An empirical analysis of backward compatibility in machine learning systems
Megha Srivastava, Besmira Nushi, Ece Kamar, Shital Shah, and Eric Horvitz · 2020
Later among the works it cites.
Positive-Congruent training: Towards Regression-Free model updates
Sijie Yan, Yuanjun Xiong, Kaustav Kundu, Shuo Yang, Siqi Deng, Meng Wang, Wei Xia, and Stefano Soatto · 2020
Later among the works it cites.
Locally adaptive label smoothing improves predictive churn
Dara Bahri and Heinrich Jiang · 2021
Later among the works it cites.
Does the whole exceed its parts? the effect of ai explanations on complementary team performance
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mariya Toneva, Alessandro Sordoni, Remi Tachet des Combes, Adam Trischler, Yoshua Bengio, and Geoffrey J Gordon · 2018
Cited alongside, same era.
FairFace: Face attribute dataset for balanced race, gender, and age
Kimmo Kärkkäinen and Jungseock Joo · 2019
Cited alongside, same era.
When does label smoothing help?
Rafael Müller, Simon Kornblith, and Geoffrey Hinton · 2019
Cited alongside, same era.
SGD on neural networks learns functions of increasing complexity
Preetum Nakkiran, Gal Kaplun, Dimitris Kalimeris, Tristan Yang, Benjamin L Edelman, Fred Zhang, and Boaz Barak · 2019
Cited alongside, same era.
A survey on image data augmentation for deep learning
Connor Shorten and Taghi M Khoshgoftaar · 2019
Cited alongside, same era.
Tesla’s “shadow” testing offers a useful advantage on the biggest problem in robocars
Brad Templeton · 2019
Cited alongside, same era.
PyHessian: Neural networks through the lens of the hessian
Zhewei Yao, Amir Gholami, Kurt Keutzer, and Michael Mahoney · 2019
Cited alongside, same era.
Gagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok, Besmira Nushi, Ece Kamar, Marco Tulio Ribeiro, and Daniel Weld · 2021
Later among the works it cites.
On the reproducibility of neural network predictions
Srinadh Bhojanapalli, Kimberly Wilber, Andreas Veit, Ankit Singh Rawat, Seungyeon Kim, Aditya Menon, and Sanjiv Kumar · 2021
Later among the works it cites.
On the importance of gradients for detecting distributional shifts in the wild
R Huang, A Geng, and Y Li · 2021
Later among the works it cites.
Churn reduction via distillation
Heinrich Jiang, Harikrishna Narasimhan, Dara Bahri, Andrew Cotter, and Afshin Rostamizadeh · 2021
Later among the works it cites.
Measuring and reducing model update regression in structured prediction for nlp
Deng Cai, Elman Mansimov, Yi-An Lai, Yixuan Su, Lei Shu, and Yi Zhang · 2022
Later among the works it cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky · 2022
Later among the works it cites.
Characterizing datapoints via Second-Split forgetting
Pratyush Maini, Saurabh Garg, Zachary C Lipton, and J Zico Kolter · 2022
Later among the works it cites.
Stacked generalization: when does it work?
Kai Ming Ting and Ian H Witten · 2022
Later among the works it cites.
ELODI: Ensemble logit difference inhibition for Positive-Congruent training
Yue Zhao, Yantao Shen, Yuanjun Xiong, Shuo Yang, Wei Xia, Zhuowen Tu, Bernt Schiele, and Stefano Soatto · 2022
Later among the works it cites.