Fetching the paper…
Reading the bibliography…
Deep neural network training often involves stochastic optimization, meaning each run will produce a different model.
A. Dvoretzky, J. Kiefer, and J. Wolfowitz, “Asymptotic minimax character of the sample distribution function and of the classical multinomial estimator,”
1956
Earlier work this paper cites.
A. Gordaliza, “Best approximations to random variables based on trimming procedures,”
1991
Earlier work this paper cites.
A. Niculescu-Mizil and R. Caruana, “Predicting good probabilities with supervised learning,” in
2005
Earlier work this paper cites.
P. C. Álvarez-Esteban, E. del Barrio, J. A. Cuesta-Albertos, and C. Matrán, “Trimmed comparison of distributions,”
2008
Earlier work this paper cites.
P. J. Huber and E. M. Ronchetti,
2009
Earlier work this paper cites.
A. Krizhevsky, “Learning multiple layers of features from tiny images,” University of Toronto, Tech. Rep., 2009. [Online]. Available:
2009
Earlier work this paper cites.
——, “Uniqueness and approximate computation of optimal incomplete transportation plans,” in
2011
Earlier work this paper cites.
F. Wei and R. M. Dudley, “Two-sample Dvoretzky–Kiefer–Wolfowitz inequalities,”
2012
Earlier work this paper cites.
P. C. Álvarez-Esteban, E. del Barrio, J. A. Cuesta-Albertos, and C. Matrán, “Similarity of samples and trimming,”
2012
Earlier work this paper cites.
M. P. Naeini, G. Cooper, and M. Hauskrecht, “Obtaining well calibrated probabilities using bayesian binning,” in
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
G. Hinton, O. Vinyals, and J. Dean, “Distilling the knowledge in a neural network,”
2015
Earlier work this paper cites.
R. Florez-Lopez and J. M. Ramon-Jeronimo, “Enhancing accuracy and interpretability of ensemble strategies in credit risk assessment. a correlated-adjusted decision forest proposal,”
2015
Cited alongside, same era.
M. Milani Fard, Q. Cormier, K. Canini, and M. Gupta, “Launch and iterate: Reducing prediction churn,” in
2016
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
B. Lakshminarayanan, A. Pritzel, and C. Blundell, “Simple and scalable predictive uncertainty estimation using deep ensembles,” in
E. del Barrio, H. Inouzhe, and C. Matrán, “Box-constrained monotone approximations to lipschitz regularizations, with applications to robust testing,”
2020
Later among the works it cites.
2021
Later among the works it cites.
X. Bouthillier, P. Delaunay, M. Bronzi, A. Trofimov, B. Nichyporuk, J. Szeto, N. Mohammadi Sepahvand, E. Raff, K. Madan, V. Voleti, S. Ebrahimi Kahou, V. Michalski, T. Arbel, C. Pal, G. Varoquaux, and P. Vincent, “Accounting for variance in machine learning benchmarks,”
2021
Later among the works it cites.
C. Summers and M. J. Dinneen, “Nondeterminism and instability in neural network optimization,”
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
P. Henderson, R. Islam, P. Bachman, J. Pineau, D. Precup, and D. Meger, “Deep reinforcement learning that matters,” in
2018
Cited alongside, same era.
A. Jacot, F. Gabriel, and C. Hongler, “Neural tangent kernel: Convergence and generalization in neural networks,”
2018
Cited alongside, same era.
X. Bouthillier, C. Laurent, and P. Vincent, “Unreproducible research is reproducible,” in
2019
Cited alongside, same era.
S. Fort, H. Hu, and B. Lakshminarayanan, “Deep ensembles: A loss landscape perspective,”
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2020
Cited alongside, same era.
E. del Barrio, H. Inouzhe, and C. Matrán, “On approximate validation of models: a Kolmogorov–Smirnov-based approach,”
2020
Cited alongside, same era.
2021
Later among the works it cites.
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo, “Swin transformer: Hierarchical vision transformer using shifted windows,” in
2021
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
D. Wood, T. Mu, A. M. Webb, H. W. Reeve, M. Lujan, and G. Brown, “A unified theory of diversity in ensemble learning,”
2023
Later among the works it cites.
2023
Later among the works it cites.
R. Theisen, H. Kim, Y. Yang, L. Hodgkinson, and M. W. Mahoney, “When are ensembles really effective?”
2023
Later among the works it cites.