Fetching the paper…
Reading the bibliography…
Machine learning algorithms often contain many hyperparameters (HPs) whose values affect the predictive performance of the induced models in intricate ways.
Kennedy J, Eberhart R (1995) Particle swarm optimization. In: Proceedings of the IEEE International Conference on Neural Networks, Perth, Australia, pp 1942 – 1948
1948
Earlier work this paper cites.
Breiman L, Friedman J, Olshen R, et al (1984) Classification and Regression Trees. Chapman & Hall (Wadsworth, Inc.)
1984
Earlier work this paper cites.
Goldberg D (1989) Genetic Algorithms in Search, Optimization and Machine Learning. Addison Wesley
1989
Earlier work this paper cites.
Quinlan JR (1993) C4.5: Programs for Machine Learning. Morgan Kaufmann Publishers Inc., San Francisco, CA, USA
1993
Earlier work this paper cites.
Kohavi R (1996) Scaling up the accuracy of naive-bayes classifiers: A decision-tree hybrid. In: Second International Conference on Knowledge Discovery and Data Mining, pp 202–207
1996
Earlier work this paper cites.
van Rijn JN, Hutter F (2017) An empirical study of hyperparameter importance across datasets. In: Proceedings of the International Workshop on Automatic Selection, Configuration and Composition of Machine Learning Algorithms co-located with the European Conference on Machine Learning & Principles and Practice of Knowledge Discovery in Databases, AutoML@PKDD/ECML 2017, Skopje, Macedonia, September 22, 2017., pp 91–98, URL http://ceur-ws.org/Vol-1998/paper_09.pdf
1998
Earlier work this paper cites.
Esposito F, Malerba D, Semeraro G, et al (1999) The effects of pruning methods on the predictive accuracy of induced decision trees. Appl Stochastic Models Bu Ind 15:277–299
1999
Earlier work this paper cites.
Liaw A, Wiener M (2002) Classification and regression by randomforest. R News 2(3):18–22
2002
Earlier work this paper cites.
Abe S (2005) Support Vector Machines for Pattern Classification. Springer London, Secaucus, NJ, USA
2005
Earlier work this paper cites.
Landwehr N, Hall M, Frank E (2005) Logistic model trees. Machine Learning 95(1-2):161–205
2005
Earlier work this paper cites.
Tan PN, Steinbach M, Kumar V (2005) Introduction to Data Mining, (First Edition), 1st edn. Addison-Wesley Longman Publishing Co., Inc., Boston, MA, USA
2005
Earlier work this paper cites.
Witten IH, Frank E (2005) Data Mining: Practical Machine Learning Tools and Techniques, 2nd edn. Morgan Kaufmann, San Francisco
2005
Earlier work this paper cites.
Ali S, Smith-Miles KA (2006) A meta-learning approach to automatic kernel selection for support vector machines. Neurocomputing 70(13):173–186
2006
Earlier work this paper cites.
Demšar J (2006) Statistical comparisons of classifiers over multiple data sets. The Journal of Machine Learning Research 7:1–30
2006
Earlier work this paper cites.
Eitrich T, Lang B (2006) Efficient optimization of support vector machine learning parameters for unbalanced datasets. Journal of Comp and Applied Mathematics 196(2):425–436
2006
Earlier work this paper cites.
Hothorn T, Hornik K, Zeileis A (2006) Unbiased recursive partitioning: A conditional inference framework. Journal of Computational and Graphical Statistics 15(3):651–674
2006
Earlier work this paper cites.
Haykin S (2007) Neural Networks: A Comprehensive Foundation (3rd Edition). Prentice-Hall, Inc., Upper Saddle River, NJ, USA
2007
Earlier work this paper cites.
Schauerhuber M, Zeileis A, Meyer D, et al (2008) Benchmarking Open-Source Tree Learners in R/RWeka, Springer Berlin Heidelberg, Berlin, Heidelberg, pp 389–396. 10.1007/978-3-540-78246-9_46
2008
Earlier work this paper cites.
Sureka A, Indukuri KV (2008) Using Genetic Algorithms for Parameter Optimization in Building Predictive Data Mining Models, Springer Berlin Heidelberg, Berlin, Heidelberg, pp 260–271. 10.1007/978-3-540-88192-6_25
2008
Earlier work this paper cites.
Brazdil P, Giraud-Carrier C, Soares C, et al (2009) Metalearning: Applications to Data Mining, 1st edn. Springer-Verlag Berlin Heidelberg
2009
Earlier work this paper cites.
Hornik K, Buchta C, Zeileis A (2009) Open-source machine learning: R meets Weka. Computational Statistics 24(2):225–232
2009
Earlier work this paper cites.
Wu X, Kumar V (2009) The Top Ten Algorithms in Data Mining, 1st edn. Chapman & Hall/CRC
2009
Earlier work this paper cites.
Ben-Hur A, Weston J (2010) A user’s guide to support vector machines. In: Data Mining Techniques for the Life Sciences, Methods in Molecular Biology, vol 609. Humana Press, p 223–239
2010
Earlier work this paper cites.
Birattari M, Yuan Z, Balaprakash P, et al (2010) F-Race and Iterated F-Race: An Overview, Springer Berlin Heidelberg, Berlin, Heidelberg, pp 311–336. 10.1007/978-3-642-02538-9_13
2010
Earlier work this paper cites.
Brodersen KH, Ong CS, Stephan KE, et al (2010) The balanced accuracy and its posterior distribution. In: Proceedings of the 2010 20th International Conference on Pattern Recognition. IEEE Computer Society, pp 3121–3124
2010
Earlier work this paper cites.
Bergstra JS, Bardenet R, Bengio Y, et al (2011) Algorithms for hyper-parameter optimization. In: Shawe-Taylor J, Zemel RS, Bartlett PL, et al (eds) Advances in Neural Information Processing Systems 24. Curran Associates, Inc., p 2546–2554
2011
Earlier work this paper cites.
Gascón-Moreno J, Salcedo-Sanz S, Ortiz-García EG, et al (2011) A binary-encoded tabu-list genetic algorithm for fast support vector regression hyper-parameters tuning. In: International Conference on Intelligent Systems Design and Applications, pp 1253–1257
2011
Earlier work this paper cites.
Hauschild M, Pelikan M (2011) An introduction and survey of estimation of distribution algorithms. Swarm and Evolutionary Computation 1(3):111 – 128
2011
Earlier work this paper cites.
Reif M, Shafait F, Dengel A (2011) Prediction of classifier training time including parameter optimization. In: Bach J, Edelkamp S (eds) KI 2011: Advances in Artificial Intelligence, Lecture Notes in Computer Science, vol 7006. Springer Berlin Heidelberg, p 260–271
2011
Earlier work this paper cites.
Barros R, Basgalupp M, de Carvalho A, et al (2012) A survey of evolutionary algorithms for decision-tree induction. Systems, Man, and Cybernetics, Part C: Applications and Reviews, IEEE Transactions on 42(3):291–312
2012
Earlier work this paper cites.
Bendtsen. C (2012) pso: Particle Swarm Optimization. URL https://CRAN.R-project.org/package=pso , r package version 1.0.3
2012
Earlier work this paper cites.
Bergstra J, Bengio Y (2012) Random search for hyper-parameter optimization. J Mach Learn Res 13:281–305
2012
Earlier work this paper cites.
Clerc M (2012) Standard partcile swarm optimization, 15 pages
2012
Earlier work this paper cites.
Gomes TAF, Prudêncio RBC, Soares C, et al (2012) Combining meta-learning and search techniques to select parameters for support vector machines. Neurocomputing 75(1):3–13
2012
Earlier work this paper cites.
Lin SW, Chen SC (2012) Parameter determination and feature selection for c4.5 algorithm using scatter search approach. Soft Computing 16(1):63–75. 10.1007/s00500-011-0734-z
2012
Earlier work this paper cites.
Ma J (2012) Parameter Tuning Using Gaussian Processes. Master’s thesis, University of Waikato, New Zealand
2012
Cited alongside, same era.
Molina MM, Luna JM, Romero C, et al (2012) Meta-learning approach for automatic parameter tuning: A case study with educational datasets. In: Proceedings of the 5th International Conference on Educational Data Mining, EDM 2012, pp 180–183
2012
Cited alongside, same era.
Reif M, Shafait F, Dengel A (2012) Meta-learning for evolutionary parameter optimization of classifiers. Machine Learning 87:357–380
2012
Cited alongside, same era.
Snoek J, Larochelle H, Adams RP (2012) Practical bayesian optimization of machine learning algorithms. In: Pereira F, Burges C, Bottou L, et al (eds) Advances in Neural Information Processing Systems 25. Curran Associates, Inc., p 2951–2959
2012
Cited alongside, same era.
Podgorelec V, Karakatic S, Barros RC, et al (2015) Evolving balanced decision trees with a multi-population genetic algorithm. In: IEEE Congress on Evolutionary Computation, CEC 2015, Sendai, Japan, May 25-28, 2015. IEEE, pp 54–61, 10.1109/CEC.2015.7256874 , URL http://ieeexplore.ieee.org/xpl/mostRecentIssue.jsp?punumber=7229815
2015
Later among the works it cites.
Sabharwal A, Samulowitz H, Tesauro G (2016) Selecting near-optimal learners via incremental data allocation. In: Proceedings of the Thirtieth AAAI Conference on Artificial Intelligence. AAAI Press, AAAI’16, pp 2007–2015, URL http://dl.acm.org/citation.cfm?id=3016100.3016179
2015
Later among the works it cites.
Therneau T, Atkinson B, Ripley B (2015) rpart: Recursive Partitioning and Regression Trees. URL https://CRAN.R-project.org/package=rpart , r package version 4.1-10
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stiglic G, Kocbek S, Pernek I, et al (2012) Comprehensive decision tree models in bioinformatics. PLOS ONE 7(3):1–13. 10.1371/journal.pone.0033812
2012
Cited alongside, same era.
Bache K, Lichman M (2013) UCI machine learning repository. URL http://archive.ics.uci.edu/ml
2013
Cited alongside, same era.
Bardenet R, Brendel M, Kégl B, et al (2013) Collaborative hyperparameter tuning. In: Dasgupta S, Mcallester D (eds) Proceedings of the 30th International Conference on Machine Learning (ICML-13), vol 28. JMLR Workshop and Conference Proceedings, pp 199–207
2013
Cited alongside, same era.
Bergstra J, Yamins D, Cox DD (2013) Making a science of model search: Hyperparameter optimization in hundreds of dimensions for vision architectures. In: Proc. 30th Intern. Conf. on Machine Learning, pp 1–9
2013
Cited alongside, same era.
Pilát M, Neruda R (2013) Multi-objectivization and Surrogate Modelling for Neural Network Hyper-parameters Tuning, Springer Berlin Heidelberg, Berlin, Heidelberg, pp 61–66. 10.1007/978-3-642-39678-6_11
2013
Cited alongside, same era.
Scrucca L (2013) Ga: A package for genetic algorithms in r. Journal of Statistical Software 53(1):1–37. 10.18637/jss.v053.i04 , URL https://www.jstatsoft.org/index.php/jss/article/view/v053i04
2013
Cited alongside, same era.
Simon D (2013) Evolutionary Optimization Algorithms, 1st edn. Wiley
2013
Cited alongside, same era.
Sun Q, Pfahringer B (2013) Pairwise meta-rules for better meta-learning-based algorithm ranking. Mach Learn 93(1):141–161. 10.1007/s10994-013-5387-y
2013
Cited alongside, same era.
Wang L, Feng M, Zhou B, et al (2015) Efficient hyper-parameter optimization for NLP applications. In: Màrquez L, Callison-Burch C, Su J, et al (eds) Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing, EMNLP 2015, Lisbon, Portugal, September 17-21, 2015. The Association for Computational Linguistics, pp 2112–2117, URL http://aclweb.org/anthology/D/D15/D15-1253.pdf
2015
Later among the works it cites.
Bischl B, Lang M, Kotthoff L, et al (2016) mlr: Machine learning in r. Journal of Machine Learning Research 17(170):1–5. URL http://jmlr.org/papers/v17/15-066.html
2016
Later among the works it cites.
European Commission (2016) Regulation (EU) 2016/679 of the European Parliament and of the Council of 27 April 2016 on the protection of natural persons with regard to the processing of personal data and on the free movement of such data, and repealing Directive 95/46/EC (General Data Protection Regulation) (Text with EEA relevance). URL https://eur-lex.europa.eu/eli/reg/2016/679/oj
2016
Later among the works it cites.
Huang BF, Boutros PC (2016) The parameter sensitivity of random forests. BMC Bioinformatics 17(1):331. 10.1186/s12859-016-1228-x , URL http://dx.doi.org/10.1186/s12859-016-1228-x
2016
Later among the works it cites.
from Jed Wing MKC, Weston S, Williams A, et al (2016) caret: Classification and Regression Training. URL https://CRAN.R-project.org/package=caret , r package version 6.0-71
2016
Later among the works it cites.
Kanda J, de Carvalho A, Hruschka E, et al (2016) Meta-learning to select the best meta-heuristic for the traveling salesman problem: A comparison of meta-features. Neurocomputing 205:393 – 406. https://doi.org/10.1016/j.neucom.2016.04.027
2016
Later among the works it cites.
Kotthoff L, Thornton C, Hoos HH, et al (2016) Auto-weka 2.0: Automatic model selection and hyperparameter optimization in weka. Journal of Machine Learning Research 17:1–5
2016
Later among the works it cites.
Lévesque JC, Gagné C, Sabourin R (2016) Bayesian hyperparameter optimization for ensemble learning. In: Proceedings of the Thirty-Second Conference on Uncertainty in Artificial Intelligence. AUAI Press, Arlington, Virginia, United States, UAI’16, pp 437–446, URL http://dl.acm.org/citation.cfm?id=3020948.3020994
2016
Later among the works it cites.
López-Ibáñez M, Dubois-Lacoste J, Cáceres LP, et al (2016) The irace package: Iterated racing for automatic algorithm configuration. Operations Research Perspectives 3:43 – 58. https://doi.org/10.1016/j.orp.2016.09.002
2016
Later among the works it cites.
Mantovani RG, Horváth T, Cerri R, et al (2016) Hyper-parameter tuning of a decision tree induction algorithm. In: 5th Brazilian Conference on Intelligent Systems, BRACIS 2016, Recife, Brazil, October 9-12, 2016. IEEE Computer Society, pp 37–42, 10.1109/BRACIS.2016.018 , URL http://ieeexplore.ieee.org/xpl/mostRecentIssue.jsp?punumber=7837801
2016
Later among the works it cites.
Massimo CM, Navarin N, Sperduti A (2016) Hyper-Parameter Tuning for Graph Kernels via Multiple Kernel Learning, Springer International Publishing, Cham, pp 214–223. 10.1007/978-3-319-46672-9_25
2016
Later among the works it cites.
Ribeiro MT, Singh S, Guestrin C (2016) Model-agnostic interpretability of machine learning. 1606.05386
2016
Later among the works it cites.
Tantithamthavorn C, McIntosh S, Hassan AE, et al (2016) Automated parameter optimization of classification techniques for defect prediction models. In: Proceedings of the 38th International Conference on Software Engineering. ACM, New York, NY, USA, ICSE ’16, pp 321–332, 10.1145/2884781.2884857
2016
Later among the works it cites.
Wainberg M, Alipanahi B, Frey BJ (2016) Are random forests truly the best classifiers? Journal of Machine Learning Research 17(110):1–5. URL http://jmlr.org/papers/v17/15-374.html
2016
Later among the works it cites.
Padierna LC, Carpio M, Rojas A, et al (2017) Hyper-Parameter Tuning for Support Vector Machines by Estimation of Distribution Algorithms, Springer International Publishing, Cham, pp 787–800
2017
Later among the works it cites.
Sanders S, Giraud-Carrier CG (2017) Informing the use of hyperparameter optimization through metalearning. In: 2017 IEEE International Conference on Data Mining, ICDM 2017, New Orleans, LA, USA, November 18-21, 2017, pp 1051–1056
2017
Later among the works it cites.
Falkner S, Klein A, Hutter F (2018) BOHB: Robust and efficient hyperparameter optimization at scale. In: Dy J, Krause A (eds) Proceedings of the 35th International Conference on Machine Learning, Proceedings of Machine Learning Research, vol 80. PMLR, pp 1437–1446
2018
Closest in time.
Garcia LPF, Lehmann J, de Carvalho ACPLF, et al (2019) New label noise injection methods for the evaluation of noise filters. Knowl Based Syst 163:693–704. 10.1016/j.knosys.2018.09.031
2018
Closest in time.
Li L, Jamieson K, DeSalvo G, et al (2018) Hyperband: A novel bandit-based approach to hyperparameter optimization. Journal of Machine Learning Research 18(185):1–52. URL http://jmlr.org/papers/v18/16-558.html
2018
Closest in time.
Blanco-Justicia A, Domingo-Ferrer J (2019) Machine learning explainability through comprehensible decision trees. In: Machine Learning and Knowledge Extraction: Third IFIP TC 5, TC 12, WG 8.4, WG 8.9, WG 12.9 International Cross-Domain Conference, CD-MAKE 2019, Canterbury, UK, August 26–29, 2019, Proceedings. Springer-Verlag, Berlin, Heidelberg, p 15–26, 10.1007/978-3-030-29726-8_2
2019
Closest in time.
Mantovani RG, Rossi AL, Alcobaça E, et al (2019) A meta-learning recommender system for hyperparameter tuning: predicting when tuning improves svm classifiers. Information Sciences 501:193–221. https://doi.org/10.1016/j.ins.2019.06.005
2019
Closest in time.
Probst P, Boulesteix A, Bischl B (2019) Tunability: Importance of hyperparameters of machine learning algorithms. J Mach Learn Res 20:53:1–53:32. URL http://jmlr.org/papers/v20/18-444.html
2019
Closest in time.
Alcobaça E, Siqueira F, Rivolli A, et al (2020) MFE: towards reproducible meta-feature extraction. J Mach Learn Res 21:111:1–111:5
2020
Closest in time.
Barella VH, Garcia LPF, de Souto MCP, et al (2021) Assessing the data complexity of imbalanced datasets. Inf Sci 553:83–109. 10.1016/j.ins.2020.12.006
2020
Closest in time.
Blanco-Justicia A, Domingo-Ferrer J, Martínez S, et al (2020) Machine learning explainability via microaggregation and shallow decision trees. Knowledge-Based Systems 194:105,532. https://doi.org/10.1016/j.knosys.2020.105532
2020
Closest in time.
Feurer M, Eggensperger K, Falkner S, et al (2020) Auto-sklearn 2.0: Hands-free automl via meta-learning. arXiv:200704074 [csLG]
2020
Closest in time.
Vieira CPR, Digiampietri LA (2020) A study about explainable articial intelligence: using decision tree to explain svm. Revista Brasileira de Computação Aplicada 12(1):113–121. 10.5335/rbca.v12i1.10247
2020
Closest in time.
2021
Closest in time.
Gijsbers P, Vanschoren J (2021) Gama: A general automated machine learning assistant. In: Dong Y, Ifrim G, Mladenić D, et al (eds) Machine Learning and Knowledge Discovery in Databases. Applied Data Science and Demo Track. Springer International Publishing, Cham, pp 560–564
2021
Closest in time.
Bischl B, Binder M, Lang M, et al (2023) Hyperparameter optimization: Foundations, algorithms, best practices and open challenges. https://wires.onlinelibrary.wiley.com/doi/10.1002/widm.1484
2023
Closest in time.
Cawley GC, Talbot NLC (2010) On over-fitting in model selection and subsequent selection bias in performance evaluation. The Journal of Machine Learning Research 11:2079–2107. URL http://www.jmlr.org/papers/v11/cawley10a.html
2079
Closest in time.