Fetching the paper…
Reading the bibliography…
When randomized ensemble methods such as bagging and random forests are implemented, a basic question arises: Is the ensemble large enough? In particular, the practitioner desires a rigorous guarantee that a given ensemble will perform nearly as well as an ideal infinite ensemble (trained on the same data).
The Annals of Mathematical Statistics , 642–669
Dvoretzky, A., Kiefer, J. and Wolfowitz, J. (1956) Asymptotic minimax character of the sample distribution function and of the classical multinomial estimator · 1956
Earlier work this paper cites.
CRC press
Breiman, L., Friedman, J., Stone, C. J. and Olshen, R. A. (1984) Classification and Regression Trees · 1984
Earlier work this paper cites.
The Annals of Probability , 234–253
Johnson, W. B., Schechtman, G. and Zinn, J. (1985) Best constants in moment inequalities for linear combinations of independent and exchangeable random variables · 1985
Earlier work this paper cites.
Journal of the American Statistical Association , 83
Bickel, P. J. and Yahav, J. A. (1988) Richardson extrapolation and the bootstrap · 1988
Earlier work this paper cites.
The Annals of Probability , 1546–1570
Talagrand, M. (1989) Isoperimetry and integrability of the sum of independent Banach-space valued random variables · 1989
Earlier work this paper cites.
The Annals of Probability , 18
Massart, P. (1990) The tight constant in the Dvoretzky-Kiefer-Wolfowitz inequality · 1990
Earlier work this paper cites.
The Annals of Probability , 19
Kwapień, S., Szulga, J. et al. (1991) Hypercontraction methods in moment inequalities for series of independent random variables in normed spaces · 1991
Earlier work this paper cites.
Machine Learning , 24
Breiman, L. (1996) Bagging predictors · 1996
Earlier work this paper cites.
Machine Learning , 45
Breiman, L. (2001) Random forests · 2001
Earlier work this paper cites.
Springer
Friedman, J., Hastie, T. and Tibshirani, R. (2001) The Elements of Statistical Learning · 2001
Earlier work this paper cites.
In Multiple Classifier Systems . Springer
Latinne, P., Debeir, O. and Decaestecker, C. (2001) Limiting the number of trees in random forests · 2001
Earlier work this paper cites.
In International Conference on Machine Learning , 377–384
Ng, A. Y. and Jordan, M. I. (2001) Convergence rates of the voting Gibbs classifier, with application to Bayesian feature selection · 2001
Earlier work this paper cites.
The Annals of Statistics , 30
Bühlmann, P. and Yu, B. (2002) Analyzing bagging · 2002
Earlier work this paper cites.
R News , 2
Liaw, A. and Wiener, M. (2002) Classification and regression by randomForest · 2002
Earlier work this paper cites.
Cambridge University Press
Sidi, A. (2003) Practical Extrapolation Methods: Theory and Applications · 2003
Earlier work this paper cites.
Journal of the Royal Statistical Society: Series B , 67
Hall, P. and Samworth, R. J. (2005) Properties of bagged nearest neighbour classifiers · 2005
Earlier work this paper cites.
BMC bioinformatics , 7
Díaz-Uriarte, R. and De Andres, S. A. (2006) Gene selection and classification of microarray data using random forest · 2006
Cited alongside, same era.
Journal of the American Statistical Association , 101
Lin, Y. and Jeon, Y. (2006) Random forests and adaptive nearest neighbors · 2006
Cited alongside, same era.
Electronic Journal of Statistics , 1
Ishwaran, H. (2007) Variable importance in binary regression trees and forests · 2007
Cited alongside, same era.
BMC bioinformatics , 9
Strobl, C., Boulesteix, A.-L., Kneib, T., Augustin, T. and Zeileis, A. (2008) Conditional variable importance for random forests · 2008
Cited alongside, same era.
Computational Statistics & Data Analysis , 53
Sexton, J. and Laake, P. (2009) Standard errors for bagged and random forest estimators · 2009
Cited alongside, same era.
The Annals of Applied Statistics , 4
Chipman, H. A., George, E. I., McCulloch, R. E. et al. (2010) Bart: Bayesian additive regression trees · 2010
Cited alongside, same era.
Journal of Machine Learning Research , 15
Wager, S., Hastie, T. and Efron, B. (2014) Confidence intervals for random forests: the jackknife and the infinitesimal jackknife · 2014
Later among the works it cites.
In 2014 IEEE International Conference on Data Mining , 1115–1120. IEEE
Zhou, F., Claire, Q. and King, R. D. (2014) Predicting the geographical origin of music · 2014
Later among the works it cites.
The R Journal , 7
Genuer, R., Poggi, J.-M. and Tuleau-Malot, C. (2015) Vsurf: an R package for variable selection using random forests · 2015
Later among the works it cites.
The Annals of Statistics , 43
Scornet, E., Biau, G. and Vert, J.-P. (2015) Consistency of random forests · 2015
Later among the works it cites.
The Journal of Machine Learning Research , 17
Blaser, R. and Fryzlewicz, P. (2016) Random rotation ensembles · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pattern Recognition Letters , 31
Genuer, R., Poggi, J.-M. and Tuleau-Malot, C. (2010) Variable selection using random forests · 2010
Cited alongside, same era.
In Data Mining (ICDM), 2011 IEEE 11th International Conference on , 41–50. IEEE
Basilico, J., Munson, M., Kolda, T., Dixon, K. and Kegelmeyer, W. (2011) Comet: A recipe for learning and using large ensembles on massive data · 2011
Cited alongside, same era.
Proteins: Structure, Function, and Bioinformatics , 79
Moult, J., Fidelis, K., Kryshtafovych, A. and Tramontano, A. (2011) Critical assessment of methods of protein structure prediction (CASP) — round IX · 2011
Cited alongside, same era.
In Computer Vision and Pattern Recognition (CVPR), 2011 IEEE Conference on , 1377–1384. IEEE
Schwing, A., Zach, C., Zheng, Y. and Pollefeys, M. (2011) Adaptive random forest – How many “experts” to ask before making a decision? · 2011
Cited alongside, same era.
Journal of Machine Learning Research , 13
Biau, G. (2012) Analysis of a random forests model · 2012
Cited alongside, same era.
In Machine Learning and Data Mining in Pattern Recognition , 154–168. Springer
Oshiro, T. M., Perez, P. S. and Baranauskas, J. A. (2012) How many trees in a random forest? · 2012
Cited alongside, same era.
Lopes, M. E. (2016) A sharp bound on the computation-accuracy tradeoff for majority voting ensembles · 2016
Later among the works it cites.
Journal of Machine Learning Research , 17
Mentch, L. and Hooker, G. (2016) Quantifying uncertainty in random forests via confidence intervals and hypothesis tests · 2016
Later among the works it cites.
Springer-Verlag New York
Wickham, H. (2016) ggplot2: Elegant Graphics for Data Analysis · 2016
Later among the works it cites.
Journal of the Royal Statistical Society Series B
Cannings, T. I. and Samworth, R. J. (2017) Random projection ensemble classification (with discussion) · 2017
Later among the works it cites.
URL http://archive.ics.uci.edu/ml
Dua, D. and Graff, C. (2017) UCI machine learning repository · 2017
Later among the works it cites.
O’Reilly Media
Géron, A. (2017) Hands-on machine learning with Scikit-Learn and TensorFlow · 2017
Later among the works it cites.
Statistics and Computing , 27
Gregorutti, B., Michel, B. and Saint-Pierre, P. (2017) Correlation and variable importance in random forests · 2017
Later among the works it cites.
Journal of Machine Learning Research , 18
Probst, P. and Boulesteix, A.-L. (2018) To tune or not to tune the number of trees in random forest · 2018
Later among the works it cites.
The Annals of Statistics , 47
Lopes, M. E. (2019) Estimating the algorithmic variance of randomized ensembles via the bootstrap · 2019
Closest in time.
Journal of Machine Learning Research , 9
Biau, G., Devroye, L. and Lugosi, G. (2008) Consistency of random forests and other averaging classifiers · 2033
Closest in time.