Fetching the paper…
Reading the bibliography…
Bayesian optimization (BO) is a powerful approach for optimizing complex and expensive-to-evaluate black-box functions.
H. Kushner, “A new method of locating the maximum point of an arbitrary multipeak curve in the presence of noise,” Journal of Basic Engineering , vol. 86, no. 1, pp. 97–106, 1964
1964
Earlier work this paper cites.
D. R. Jones, M. Schonlau, and W. J. Welch, “Efficient global optimization of expensive black-box functions,” Journal of Global optimization , vol. 13, pp. 455–492, 1998
1998
Earlier work this paper cites.
C. E. Rasmussen, C. K. Williams et al. , Gaussian processes for machine learning . Springer, 2006, vol. 1
2006
Earlier work this paper cites.
D. J. Lizotte, T. Wang, M. H. Bowling, D. Schuurmans et al. , “Automatic gait optimization with gaussian process regression.” in IJCAI , vol. 7, 2007, pp. 944–949
2007
Earlier work this paper cites.
L. Van der Maaten and G. Hinton, “Visualizing data using t-sne.” Journal of machine learning research , vol. 9, no. 11, 2008
2008
Earlier work this paper cites.
E. Brochu, T. Brochu, and N. De Freitas, “A bayesian interactive optimization approach to procedural animation design,” in Proceedings of the 2010 ACM SIGGRAPH/Eurographics Symposium on Computer Animation , 2010, pp. 103–112
2010
Earlier work this paper cites.
N. Srinivas, A. Krause, S. Kakade, and M. Seeger, “Gaussian process optimization in the bandit setting: no regret and experimental design,” in Proceedings of the 27th International Conference on International Conference on Machine Learning , 2010, pp. 1015–1022
2010
Earlier work this paper cites.
J. Bergstra, R. Bardenet, Y. Bengio, and B. Kégl, “Algorithms for hyper-parameter optimization,” Advances in neural information processing systems , vol. 24, 2011
2011
Earlier work this paper cites.
F. Hutter, H. H. Hoos, and K. Leyton-Brown, “Sequential model-based optimization for general algorithm configuration,” in International Conference on Learning and Intelligent Optimization . Springer, 2011, pp. 507–523
2011
Earlier work this paper cites.
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay, “Scikit-learn: Machine learning in Python,” Journal of Machine Learning Research , vol. 12, pp. 2825–2830, 2011. [Online]. Available: http://www.jmlr.org/papers/volume12/pedregosa11a/pedregosa11a.pdf
2011
Earlier work this paper cites.
J. Snoek, H. Larochelle, and R. P. Adams, “Practical bayesian optimization of machine learning algorithms,” Advances in neural information processing systems , vol. 25, 2012
2012
Earlier work this paper cites.
P. D. Arendt, D. W. Apley, and W. Chen, “Quantification of model uncertainty: Calibration, model discrepancy, and identifiability,” Journal of mechanical design , vol. 134, no. 10, 2012
2012
Earlier work this paper cites.
M. Sugiyama, T. Suzuki, and T. Kanamori, Density ratio estimation in machine learning . Cambridge University Press, 2012
2012
Earlier work this paper cites.
J. O. Berger, Statistical decision theory and Bayesian analysis . Springer Science & Business Media, 2013
2013
Earlier work this paper cites.
K. Swersky, J. Snoek, and R. P. Adams, “Multi-task bayesian optimization,” Advances in neural information processing systems , vol. 26, 2013
2013
Earlier work this paper cites.
R. Bardenet, M. Brendel, B. Kégl, and M. Sebag, “Collaborative hyperparameter tuning,” in International conference on machine learning . PMLR, 2013, pp. 199–207
2013
Earlier work this paper cites.
S. J. Pocock, C. A. Ariti, J. J. McMurray, A. Maggioni, L. Køber, I. B. Squire, K. Swedberg, J. Dobson, K. K. Poppe, G. A. Whalley et al. , “Predicting survival in heart failure: a risk score based on 39 372 patients from 30 studies,” European Heart Journal , vol. 34, no. 19, pp. 1404–1413, 2013
2013
Earlier work this paper cites.
T. Desautels, A. Krause, and J. W. Burdick, “Parallelizing exploration-exploitation tradeoffs in gaussian process bandit optimization,” Journal of Machine Learning Research , vol. 15, pp. 3873–3923, 2014
2014
Earlier work this paper cites.
D. Duvenaud, “Automatic model construction with gaussian processes,” Ph.D. dissertation, University of Cambridge, 2014
2014
Earlier work this paper cites.
J. Snoek, O. Rippel, K. Swersky, R. Kiros, N. Satish, N. Sundaram, M. Patwary, Prabhat, and R. P. Adams, “Scalable bayesian optimization using deep neural networks,” in International Conference on Machine Learning , 2015, pp. 2171–2180
2015
Earlier work this paper cites.
J. Bergstra, B. Komer, C. Eliasmith, D. Yamins, and D. D. Cox, “Hyperopt: A python library for model selection and hyperparameter optimization,” Computational Science & Discovery , vol. 8, no. 1, p. 014008, 2015
2015
Earlier work this paper cites.
M. Feurer, J. T. Springenberg, and F. Hutter, “Initializing bayesian hyper-parameter optimization via meta-learning,” in Twenty-Ninth AAAI Conference on Artificial Intelligence , 2015
2015
Earlier work this paper cites.
R. Calandra, A. Seyfarth, J. Peters, and M. P. Deisenroth, “Bayesian optimization for learning gaits under uncertainty: An experimental comparison on a dynamic bipedal walker,” Annals of Mathematics and Artificial Intelligence , vol. 76, pp. 5–23, 2016
2016
Earlier work this paper cites.
A. G. Wilson, Z. Hu, R. Salakhutdinov, and E. P. Xing, “Deep kernel learning,” in Proceedings of the 19th International Conference on Artificial Intelligence and Statistics , ser. Proceedings of Machine Learning Research, A. Gretton and C. C. Robert, Eds., vol. 51. Cadiz, Spain: PMLR, 09–11 May 2016, pp. 370–378. [Online]. Available: https://proceedings.mlr.press/v51/wilson16.html
2016
Earlier work this paper cites.
R. Calandra, J. Peters, C. E. Rasmussen, and M. P. Deisenroth, “Manifold gaussian processes for regression,” in 2016 International joint conference on neural networks (IJCNN) . IEEE, 2016, pp. 3338–3345
2016
Earlier work this paper cites.
J. T. Springenberg, A. Klein, S. Falkner, and F. Hutter, “Bayesian optimization with robust bayesian neural networks,” in Advances in Neural Information Processing Systems , D. Lee, M. Sugiyama, U. Luxburg, I. Guyon, and R. Garnett, Eds., vol. 29. Curran Associates, Inc., 2016. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2016/file/a96d3afec184766bfeca7a9f989fc7e7-Paper.pdf
2016
Earlier work this paper cites.
M. Poloczek, J. Wang, and P. I. Frazier, “Warm starting bayesian optimization,” in 2016 Winter Simulation Conference (WSC) . IEEE, 2016, pp. 770–781
2016
Cited alongside, same era.
M. A. Duggan, W. F. Anderson, S. Altekruse, L. Penberthy, and M. E. Sherman, “The surveillance, epidemiology and end results (SEER) program and pathology: towards strengthening the critical relationship,” The American Journal of Surgical Pathology , vol. 40, no. 12, p. e94, 2016
2016
Cited alongside, same era.
P. I. Frazier, “A tutorial on bayesian optimization,” arXiv preprint arXiv:1807.02811 , 2018
2018
Cited alongside, same era.
A. Hebbal, L. Brevault, M. Balesdent, E.-G. Taibi, and N. Melab, “Efficient global optimization using deep gaussian processes,” in 2018 IEEE Congress on evolutionary computation (CEC) . IEEE, 2018, pp. 1–8
2018
Cited alongside, same era.
S. M. Xie, A. Raghunathan, P. Liang, and T. Ma, “An explanation of in-context learning as implicit bayesian inference,” in International Conference on Learning Representations , 2022. [Online]. Available: https://openreview.net/forum?id=RdJVFCHjUMI
2022
Later among the works it cites.
Z. Dai, Y. Shu, B. K. H. Low, and P. Jaillet, “Sample-then-optimize batch neural thompson sampling,” Advances in Neural Information Processing Systems , vol. 35, pp. 23 331–23 344, 2022
2022
Later among the works it cites.
Y. Chen, X. Song, C. Lee, Z. Wang, R. Zhang, D. Dohan, K. Kawakami, G. Kochanski, A. Doucet, M. Ranzato et al. , “Towards learning universal hyperparameter optimizers with transformers,” Advances in Neural Information Processing Systems , vol. 35, pp. 32 053–32 068, 2022
2022
Later among the works it cites.
D. Jarrett, B. C. Cebere, T. Liu, A. Curth, and M. van der Schaar, “Hyperimpute: Generalized iterative imputation with automatic model selection,” in International Conference on Machine Learning . PMLR, 2022, pp. 9916–9937
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Falkner, A. Klein, and F. Hutter, “Bohb: Robust and efficient hyperparameter optimization at scale,” in International conference on machine learning . PMLR, 2018, pp. 1437–1446
2018
Cited alongside, same era.
2018
Cited alongside, same era.
K. Kandasamy, W. Neiswanger, J. Schneider, B. Poczos, and E. P. Xing, “Neural architecture search with bayesian optimisation and optimal transport,” Advances in neural information processing systems , vol. 31, 2018
2018
Cited alongside, same era.
T. Akiba, S. Sano, T. Yanase, T. Ohta, and M. Koyama, “Optuna: A next-generation hyperparameter optimization framework,” in Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining , 2019
2019
Cited alongside, same era.
P. C. U. PCUK, “Cutract,” https://prostatecanceruk.org , 2019
2019
Cited alongside, same era.
D. Eriksson, M. Pearce, J. Gardner, R. D. Turner, and M. Poloczek, “Scalable global optimization via local Bayesian optimization,” in Advances in Neural Information Processing Systems , 2019, pp. 5496–5507. [Online]. Available: http://papers.nips.cc/paper/8788-scalable-global-optimization-via-local-bayesian-optimization.pdf
2019
Cited alongside, same era.
S. Greenhill, S. Rana, S. Gupta, P. Vellanki, and S. Venkatesh, “Bayesian optimization for adaptive experimental design: A review,” IEEE access , vol. 8, pp. 13 937–13 948, 2020
2020
Cited alongside, same era.
K. Korovina, S. Xu, K. Kandasamy, W. Neiswanger, B. Poczos, J. Schneider, and E. Xing, “Chembo: Bayesian optimization of small organic molecules with synthesizable recommendations,” in International Conference on Artificial Intelligence and Statistics . PMLR, 2020, pp. 3393–3403
2020
Cited alongside, same era.
2022
Later among the works it cites.
T. Dinh, Y. Zeng, R. Zhang, Z. Lin, M. Gira, S. Rajput, J. Y. Sohn, D. Papailiopoulos, and K. Lee, “Lift: Language-interfaced fine-tuning for non-language machine learning tasks,” in Advances in Neural Information Processing Systems , A. H. Oh, A. Agarwal, D. Belgrave, and K. Cho, Eds., 2022
2022
Later among the works it cites.
Y. Lu, M. Bartolo, A. Moore, S. Riedel, and P. Stenetorp, “Fantastically ordered prompts and where to find them: Overcoming few-shot prompt order sensitivity,” in Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2022, pp. 8086–8098
2022
Later among the works it cites.
A. I. Cowen-Rivers, W. Lyu, R. Tutunov, Z. Wang, A. Grosnit, R. R. Griffiths, A. M. Maraval, H. Jianye, J. Wang, J. Peters et al. , “Hebo: Pushing the limits of sample-efficient hyper-parameter optimisation,” Journal of Artificial Intelligence Research , vol. 74, pp. 1269–1349, 2022
2022
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Holt, M. R. Luyten, and M. van der Schaar, “L2mac: Large language model automatic computer for unbounded code generation,” in The Twelfth International Conference on Learning Representations , 2023
2023
Later among the works it cites.
K. Singhal, S. Azizi, T. Tu, S. S. Mahdavi, J. Wei, H. W. Chung, N. Scales, A. Tanwani, H. Cole-Lewis, S. Pfohl et al. , “Large language models encode clinical knowledge,” Nature , pp. 1–9, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
F. Imrie, B. Cebere, E. F. McKinney, and M. van der Schaar, “Autoprognosis 2.0: Democratizing diagnostic and prognostic modeling in healthcare with automated machine learning,” PLOS Digital Health , vol. 2, no. 6, p. e0000276, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
J. Lehman, J. Gordon, S. Jain, K. Ndousse, C. Yeh, and K. O. Stanley, “Evolution through large models,” in Handbook of Evolutionary Machine Learning . Springer, 2023, pp. 331–366
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Hegselmann, A. Buendia, H. Lang, M. Agrawal, X. Jiang, and D. Sontag, “Tabllm: Few-shot classification of tabular data with large language models,” in International Conference on Artificial Intelligence and Statistics . PMLR, 2023, pp. 5549–5581
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Sun, A. Hüyük, and M. van der Schaar, “Query-dependent prompt evaluation and optimization with offline inverse RL,” in The Twelfth International Conference on Learning Representations , 2024. [Online]. Available: https://openreview.net/forum?id=N6o0ZtPzTg
2024
Closest in time.
Q. Guo, R. Wang, J. Guo, B. Li, K. Song, X. Tan, G. Liu, J. Bian, and Y. Yang, “Connecting large language models with evolutionary algorithms yields powerful prompt optimizers,” in The Twelfth International Conference on Learning Representations , 2024. [Online]. Available: https://openreview.net/forum?id=ZG3RaNIsO8
2024
Closest in time.
A. Chen, D. Dohan, and D. So, “Evoprompting: Language models for code-level neural architecture search,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.