Fetching the paper…
Reading the bibliography…
The ever-growing demand and complexity of machine learning are putting pressure on hyper-parameter tuning systems: while the evaluation cost of models continues to increase, the scalability of state-of-the-arts starts to become a crucial bottleneck.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Efficient global optimization of expensive black-box functions
Donald R Jones, Matthias Schonlau, and William J Welch. 1998 · 1998
Earlier work this paper cites.
Model-based Asynchronous Hyperparameter Optimization
Louis C. Tiao, Aaron Klein, Cédric Archambeau, and Matthias W. Seeger. 2020 · 2003
Earlier work this paper cites.
Batch bayesian optimization via simulation matching. In Advances in Neural Information Processing Systems . 109–117
Javad Azimi, Alan Fern, and Xiaoli Z Fern. 2010 · 2010
Earlier work this paper cites.
Gaussian Process Optimization in the Bandit Setting: No Regret and Experimental Design. In Proceedings of the 27th International Conference on Machine Learning . Omnipress
Niranjan Srinivas, Andreas Krause, Sham Kakade, and Matthias Seeger. 2010 · 2010
Earlier work this paper cites.
Algorithms for hyper-parameter optimization. In Advances in neural information processing systems . 2546–2554
James S Bergstra, Rémi Bardenet, Yoshua Bengio, and Balázs Kégl. 2011 · 2011
Earlier work this paper cites.
SystemML: Declarative machine learning on MapReduce. In 2011 IEEE 27th International Conference on Data Engineering . IEEE, 231–242
Amol Ghoting, Rajasekar Krishnamurthy, Edwin Pednault, Berthold Reinwald, Vikas Sindhwani, Shirish Tatikonda, Yuanyuan Tian, and Shivakumar Vaithyanathan. 2011 · 2011
Earlier work this paper cites.
Sequential model-based optimization for general algorithm configuration. In International Conference on Learning and Intelligent Optimization . Springer, 507–523
Frank Hutter, Holger H Hoos, and Kevin Leyton-Brown. 2011 · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio. 2012 · 2012
Earlier work this paper cites.
Practical bayesian optimization of machine learning algorithms. In Advances in neural information processing systems
Jasper Snoek, Hugo Larochelle, and Ryan P Adams. 2012 · 2012
Earlier work this paper cites.
Bayesian optimization in high dimensions via random embeddings. In Twenty-Third International Joint Conference on Artificial Intelligence
Ziyu Wang, Masrour Zoghi, Frank Hutter, David Matheson, and Nando De Freitas. 2013 · 2013
Earlier work this paper cites.
OpenML: networked science in machine learning
Joaquin Vanschoren, Jan N Van Rijn, Bernd Bischl, and Luis Torgo. 2014 · 2014
Earlier work this paper cites.
Speeding Up Automatic Hyperparameter Optimization of Deep Neural Networks by Extrapolation of Learning Curves.. In IJCAI , Vol. 15. 3460–8
Tobias Domhan, Jost Tobias Springenberg, and Frank Hutter. 2015 · 2015
Earlier work this paper cites.
Efficient and robust automated machine learning. In Advances in neural information processing systems . 2962–2970
Matthias Feurer, Aaron Klein, Katharina Eggensperger, Jost Springenberg, Manuel Blum, and Frank Hutter. 2015 · 2015
Earlier work this paper cites.
Xgboost: A scalable tree boosting system. In Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 785–794
Tianqi Chen and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Batch bayesian optimization via local penalization. In Artificial Intelligence and Statistics . 648–657
Javier González, Zhenwen Dai, Philipp Hennig, and Neil Lawrence. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Non-stochastic best arm identification and hyperparameter optimization. In Artificial Intelligence and Statistics . 240–248
Kevin Jamieson and Ameet Talwalkar. 2016 · 2016
Earlier work this paper cites.
ModelDB: a system for machine learning model management. In Proceedings of the Workshop on Human-In-the-Loop Data Analytics . 1–3
Manasi Vartak, Harihar Subramanyam, Wei-En Lee, Srinidhi Viswanathan, Saadiyah Husnoo, Samuel Madden, and Matei Zaharia. 2016 · 2016
Earlier work this paper cites.
Practical neural network performance prediction for early stopping
Bowen Baker, Otkrist Gupta, Ramesh Raskar, and Nikhil Naik. 2017 · 2017
Earlier work this paper cites.
Tfx: A tensorflow-based production-scale machine learning platform. In Proceedings of the 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . 1387–1395
Denis Baylor, Eric Breck, Heng-Tze Cheng, Noah Fiedel, Chuan Yu Foo, Zakaria Haque, Salem Haykal, Mustafa Ispir, Vihan Jain, Levent Koc, et al · 2017
Earlier work this paper cites.
Google vizier: A service for black-box optimization. In Proceedings of the 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 1487–1495
Daniel Golovin, Benjamin Solnik, Subhodeep Moitra, Greg Kochanski, John Karro, and D Sculley. 2017 · 2017
Earlier work this paper cites.
Population based training of neural networks
Max Jaderberg, Valentin Dalibard, Simon Osindero, Wojciech M Czarnecki, Jeff Donahue, Ali Razavi, Oriol Vinyals, Tim Green, Iain Dunning, Karen Simonyan, et al · 2017
Earlier work this paper cites.
Multi-fidelity bayesian optimisation with continuous approximations
Kirthevasan Kandasamy, Gautam Dasarathy, Jeff Schneider, and Barnabas Poczos. 2017a · 2017
Cited alongside, same era.
Asynchronous Parallel Bayesian Optimisation via Thompson Sampling
Kirthevasan Kandasamy, Akshay Krishnamurthy, Jeff Schneider, and Barnabás Póczos. 2017b · 2017
Cited alongside, same era.
Learning curve prediction with Bayesian neural networks
Aaron Klein, Stefan Falkner, Jost Tobias Springenberg, and Frank Hutter. 2017c · 2017
Cited alongside, same era.
Multi-information source optimization. In Advances in Neural Information Processing Systems . 4288–4298
Matthias Poloczek, Jialei Wang, and Peter Frazier. 2017 · 2017
Cited alongside, same era.
HoloClean: Holistic Data Repairs with Probabilistic Inference
Theodoros Rekatsinas, Xu Chu, Ihab F Ilyas, and Christopher Ré. 2017 · 2017
Cited alongside, same era.
Automated machine learning: methods, systems, challenges
Frank Hutter, Lars Kotthoff, and Joaquin Vanschoren. 2019 · 2019
Later among the works it cites.
Qtune: A query-aware database tuning system with deep reinforcement learning
Guoliang Li, Xuanhe Zhou, Shifu Li, and Bo Gao. 2019 · 2019
Later among the works it cites.
Deep neural architecture search with deep graph bayesian optimization. In 2019 IEEE/WIC/ACM International Conference on Web Intelligence (WI) . IEEE, 500–507
Lizheng Ma, Jiaxu Cui, and Bo Yang. 2019 · 2019
Later among the works it cites.
Incremental and approximate inference for faster occlusion-based deep cnn explanations. In Proceedings of the 2019 International Conference on Management of Data . 1589–1606
Supun Nakandala, Arun Kumar, and Yannis Papakonstantinou. 2019 · 2019
Later among the works it cites.
TPOT: A tree-based pipeline optimization tool for automating machine learning
Randal S Olson and Jason H Moore. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
BOHB: Robust and efficient hyperparameter optimization at scale. In International Conference on Machine Learning . PMLR, 1437–1446
Stefan Falkner, Aaron Klein, and Frank Hutter. 2018 · 2018
Cited alongside, same era.
Automated Machine Learning: Methods, Systems, Challenges
Frank Hutter, Lars Kotthoff, and Joaquin Vanschoren (Eds.). 2018 · 2018
Cited alongside, same era.
Neural Architecture Search with Bayesian Optimisation and Optimal Transport
Kirthevasan Kandasamy, Willie Neiswanger, Jeff Schneider, Barnabas Poczos, and Eric P Xing. 2018b · 2018
Cited alongside, same era.
Northstar: An Interactive Data Science System
Tim Kraska. 2018 · 2018
Cited alongside, same era.
Hyperband: A novel bandit-based approach to hyperparameter optimization
Lisha Li, Kevin Jamieson, Giulia DeSalvo, Afshin Rostamizadeh, and Ameet Talwalkar. 2018a · 2018
Cited alongside, same era.
Massively Parallel Hyperparameter Tuning
Liam Li, Kevin Jamieson, Afshin Rostamizadeh, Ekaterina Gonina, Moritz Hardt, Benjamin Recht, and Ameet Talwalkar. 2018b · 2018
Cited alongside, same era.
Tune: A Research Platform for Distributed Model Selection and Training
Richard Liaw, Eric Liang, Robert Nishihara, Philipp Moritz, Joseph E Gonzalez, and Ion Stoica. 2018 · 2018
Cited alongside, same era.
Regularized evolution for image classifier architecture search. In Proceedings of the AAAI conference on artificial intelligence , Vol. 33. 4780–4789
Esteban Real, Alok Aggarwal, Yanping Huang, and Quoc V Le. 2019 · 2019
Later among the works it cites.
Democratizing data science through interactive curation of ml pipelines. In Proceedings of the 2019 International Conference on Management of Data . 1171–1188
Zeyuan Shang, Emanuel Zgraggen, Benedetto Buratti, Ferdinand Kossmann, Philipp Eichmann, Yeounoh Chung, Carsten Binnig, Eli Upfal, and Tim Kraska. 2019 · 2019
Later among the works it cites.
Bananas: Bayesian optimization with neural architectures for neural architecture search
Colin White, Willie Neiswanger, and Yash Savani. 2019 · 2019
Later among the works it cites.
Practical multi-fidelity Bayesian optimization for hyperparameter tuning
Jian Wu, Saul Toscanopalmerin, Peter I Frazier, and Andrew Gordon Wilson. 2019 · 2019
Later among the works it cites.
PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture Search. In International Conference on Learning Representations
Yuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen, Guo-Jun Qi, Qi Tian, and Hongkai Xiong. 2019 · 2019
Later among the works it cites.
Differential Evolution for Neural Architecture Search
Noor Awad, Neeratyoy Mallik, and Frank Hutter. 2020 · 2020
Later among the works it cites.
SystemDS: A Declarative Machine Learning System for the End-to-End Data Science Lifecycle. In Conference on Innovative Data Systems Research
Matthias Boehm, Iulian Antonov, Sebastian Baunsgaard, Mark Dokter, Robert Erich Ginthoer, Kevin Innerebner, Florijan Klezin, Stefanie Lindstaedt, Arnab Phani, Benjamin Rath, et al · 2020
Later among the works it cites.
Language models are few-shot learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Later among the works it cites.
A System for Massively Parallel Hyperparameter Tuning
Liam Li, Kevin Jamieson, Afshin Rostamizadeh, Ekaterina Gonina, Jonathan Ben-tzur, Moritz Hardt, Benjamin Recht, and Ameet Talwalkar. 2020a · 2020
Later among the works it cites.
Cerebro: A data system for optimized deep learning model selection
Supun Nakandala, Yuhao Zhang, and Arun Kumar. 2020 · 2020
Later among the works it cites.
Snorkel: Rapid training data creation with weak supervision
Alexander Ratner, Stephen H Bach, Henry Ehrenberg, Jason Fries, Sen Wu, and Christopher Ré. 2020 · 2020
Later among the works it cites.
Neural architecture search using bayesian optimisation with weisfeiler-lehman kernel
Binxin Ru, Xingchen Wan, Xiaowen Dong, and Michael Osborne. 2020 · 2020
Later among the works it cites.
NAS-Bench-301 and the case for surrogate benchmarks for neural architecture search
Julien Siems, Lucas Zimmer, Arber Zela, Jovita Lukasik, Margret Keuper, and Frank Hutter. 2020 · 2020
Later among the works it cites.
Zeroer: Entity resolution using zero labeled examples. In Proceedings of the 2020 ACM SIGMOD International Conference on Management of Data . 1149–1164
Renzhi Wu, Sanya Chaba, Saurabh Sawlani, Xu Chu, and Saravanan Thirumuruganathan. 2020a · 2020
Later among the works it cites.
Complaint-driven training data debugging for query 2.0. In Proceedings of the 2020 ACM SIGMOD International Conference on Management of Data . 1317–1334
Weiyuan Wu, Lampros Flokas, Eugene Wu, and Jiannan Wang. 2020b · 2020
Later among the works it cites.
OpenBox: A Generalized Black-box Optimization Service
Yang Li, Yu Shen, Wentao Zhang, Yuanwei Chen, Huaijun Jiang, Mingchao Liu, Jiawei Jiang, Jinyang Gao, Wentao Wu, Zhi Yang, Ce Zhang, and Bin Cui. 2021b · 2021
Later among the works it cites.
VolcanoML: Speeding up End-to-End AutoML via Scalable Search Space Decomposition
Yang Li, Yu Shen, Wentao Zhang, Jiawei Jiang, Bolin Ding, Yaliang Li, Jingren Zhou, Zhi Yang, Wentao Wu, Ce Zhang, and Bin Cui. 2021c · 2021
Later among the works it cites.
ResTune: Resource Oriented Tuning Boosted by Meta-Learning for Cloud Databases. In Proceedings of the 2021 International Conference on Management of Data . 2102–2114
Xinyi Zhang, Hong Wu, Zhuo Chang, Shuowei Jin, Jian Tan, Feifei Li, Tieying Zhang, and Bin Cui. 2021 · 2021
Later among the works it cites.