Fetching the paper…
Reading the bibliography…
Meta-learning, or learning to learn, is a machine learning approach that utilizes prior learning experiences to expedite the learning process on unseen tasks.
Srivastava N, Hinton GE, Krizhevsky A, Sutskever I, Salakhutdinov R (2014) Dropout: a simple way to prevent neural networks from overfitting. Journal of Machine Learning Research 15(1):1929–1958
1958
Earlier work this paper cites.
Kuurkova V (1991) Kolmogorov’s theorem is relevant. Neural computation 3(4):617–622
1991
Earlier work this paper cites.
Jones DR, Schonlau M, Welch WJ (1998) Efficient global optimization of expensive black-box functions. J Global Optimization 13(4):455–492
1998
Earlier work this paper cites.
Breiman L (2001) Random forests. Machine learning 45(1):5–32
2001
Earlier work this paper cites.
Borg I, Groenen P (2003) Modern multidimensional scaling: Theory and applications. Journal of Educational Measurement 40(3):277–280
2003
Earlier work this paper cites.
Rasmussen CE (2003) Gaussian processes in machine learning. In: Summer School on Machine Learning, Springer, pp 63–71
2003
Earlier work this paper cites.
Castiello C, Castellano G, Fanelli AM (2005) Meta-data: Characterization of input features for meta-learning. In: MDAI, Springer, Lecture Notes in Computer Science, vol 3558, pp 457–468
2005
Earlier work this paper cites.
Demšar J (2006) Statistical comparisons of classifiers over multiple data sets. Journal of Machine learning research 7(Jan):1–30
2006
Earlier work this paper cites.
Segrera S, Lucas JP, García MNM (2008) Information-theoretic measures for meta-learning. In: HAIS, Springer, Lecture Notes in Computer Science, vol 5271, pp 458–465
2008
Earlier work this paper cites.
Bottou L (2010) Large-scale machine learning with stochastic gradient descent. In: Proceedings of COMPSTAT’2010, Springer, pp 177–186
2010
Earlier work this paper cites.
Bergstra JS, Bardenet R, Bengio Y, Kégl B (2011) Algorithms for hyper-parameter optimization. In: Advances in neural information processing systems, pp 2546–2554
2011
Earlier work this paper cites.
Hutter F, Hoos HH, Leyton-Brown K (2011) Sequential model-based optimization for general algorithm configuration. In: International conference on learning and intelligent optimization, Springer, pp 507–523
2011
Earlier work this paper cites.
Pedregosa F, Varoquaux G, Gramfort A, Michel V, Thirion B, Grisel O, Blondel M, Prettenhofer P, Weiss R, Dubourg V, Vanderplas J, Passos A, Cournapeau D, Brucher M, Perrot M, Duchesnay E (2011) Scikit-learn: Machine learning in Python. Journal of Machine Learning Research 12:2825–2830
2011
Earlier work this paper cites.
Bergstra J, Bengio Y (2012) Random search for hyper-parameter optimization. Journal of machine learning research 13(Feb):281–305
2012
Earlier work this paper cites.
Tieleman T, Hinton G (2012) Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude. COURSERA: Neural networks for machine learning 4(2):26–31
2012
Earlier work this paper cites.
Bardenet R, Brendel M, Kégl B, Sebag M (2013) Collaborative hyperparameter tuning. In: International conference on machine learning, pp 199–207
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
Reif M, Shafait F, Goldstein M, Breuel TM, Dengel A (2014) Automatic classifier selection for non-experts. Pattern Anal Appl 17(1):83–96
2014
Cited alongside, same era.
Vanschoren J, Van Rijn JN, Bischl B, Torgo L (2014) Openml: networked science in machine learning. ACM SIGKDD Explorations Newsletter 15(2):49–60
2014
Cited alongside, same era.
Yogatama D, Mann G (2014) Efficient transfer learning method for automatic hyperparameter tuning. In: AISTATS, JMLR.org, JMLR Workshop and Conference Proceedings, vol 33, pp 1077–1085
2014
Cited alongside, same era.
Feurer M, Springenberg JT, Hutter F (2015) Initializing bayesian hyperparameter optimization via meta-learning. In: Bonet B, Koenig S (eds) Proceedings of the Twenty-Ninth AAAI Conference on Artificial Intelligence, January 25-30, 2015, Austin, Texas, USA, AAAI Press, pp 1128–1135, URL http://www.aaai.org/ocs/index.php/AAAI/AAAI15/paper/view/10029
2015
Cited alongside, same era.
Snell J, Swersky K, Zemel R (2017) Prototypical networks for few-shot learning. In: Advances in neural information processing systems, pp 4077–4087
2017
Later among the works it cites.
Zaheer M, Kottur S, Ravanbakhsh S, Póczos B, Salakhutdinov RR, Smola AJ (2017) Deep sets. In: NIPS, pp 3394–3404
2017
Later among the works it cites.
Berlemont S, Lefebvre G, Duffner S, Garcia C (2018) Class-balanced siamese neural networks. Neurocomputing 273:47–56
2018
Later among the works it cites.
Falkner S, Klein A, Hutter F (2018) BOHB: robust and efficient hyperparameter optimization at scale 80:1436–1445, URL http://proceedings.mlr.press/v80/falkner18a.html
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Filchenkov A, Pendryak A (2015) Datasets meta-feature description for recommending feature selection algorithm. In: 2015 Artificial Intelligence and Natural Language and Information Extraction, Social Media and Web Search FRUCT Conference (AINL-ISMW FRUCT), IEEE, pp 11–18
2015
Cited alongside, same era.
Ioffe S, Szegedy C (2015) Batch normalization: Accelerating deep network training by reducing internal covariate shift. In: ICML, JMLR.org, JMLR Workshop and Conference Proceedings, vol 37, pp 448–456
2015
Cited alongside, same era.
2015
Cited alongside, same era.
Koch G, Zemel R, Salakhutdinov R (2015) Siamese neural networks for one-shot image recognition. In: ICML deep learning workshop, vol 2
2015
Cited alongside, same era.
Wistuba M, Schilling N, Schmidt-Thieme L (2015) Sequential model-free hyperparameter tuning. In: ICDM, IEEE Computer Society, pp 1033–1038
2015
Cited alongside, same era.
Abadi M, Barham P, Chen J, Chen Z, Davis A, Dean J, Devin M, Ghemawat S, Irving G, Isard M, Kudlur M, Levenberg J, Monga R, Moore S, Murray DG, Steiner B, Tucker PA, Vasudevan V, Warden P, Wicke M, Yu Y, Zheng X (2016) Tensorflow: A system for large-scale machine learning. In: OSDI, USENIX Association, pp 265–283
2016
Cited alongside, same era.
Song HO, Xiang Y, Jegelka S, Savarese S (2016) Deep metric learning via lifted structured feature embedding. In: CVPR, IEEE Computer Society, pp 4004–4012
2016
Cited alongside, same era.
Springenberg JT, Klein A, Falkner S, Hutter F (2016) Bayesian optimization with robust bayesian neural networks. In: Advances in neural information processing systems, pp 4134–4142
2016
Cited alongside, same era.
2018
Later among the works it cites.
Finn C, Xu K, Levine S (2018) Probabilistic model-agnostic meta-learning. In: NeurIPS, pp 9537–9548
2018
Later among the works it cites.
Hewitt LB, Nye MI, Gane A, Jaakkola TS, Tenenbaum JB (2018) The variational homoencoder: Learning to learn high capacity generative models from few examples. In: UAI, AUAI Press, pp 988–997
2018
Later among the works it cites.
Lindauer M, Hutter F (2018) Warmstarting of model-based algorithm configuration. In: AAAI, AAAI Press, pp 1355–1362
2018
Later among the works it cites.
Perrone V, Jenatton R, Seeger MW, Archambeau C (2018) Scalable hyperparameter transfer learning. In: NeurIPS, pp 6846–6856
2018
Later among the works it cites.
2018
Later among the works it cites.
Vanschoren J (2018) Meta-learning: A survey. arXiv preprint arXiv:181003548
2018
Later among the works it cites.
Wistuba M, Schilling N, Schmidt-Thieme L (2018) Scalable gaussian process-based transfer surrogates for hyperparameter optimization. Machine Learning 107(1):43–78
2018
Later among the works it cites.
Yoon J, Kim T, Dia O, Kim S, Bengio Y, Ahn S (2018) Bayesian model-agnostic meta-learning. In: NeurIPS, pp 7343–7353
2018
Later among the works it cites.
Zheng Z, Zheng L, Yang Y (2018) A discriminatively learned CNN embedding for person reidentification. TOMCCAP 14(1):13:1–13:20
2018
Later among the works it cites.
Achille A, Lam M, Tewari R, Ravichandran A, Maji S, Fowlkes C, Soatto S, Perona P (2019) Task2vec: Task embedding for meta-learning. CoRR
2019
Closest in time.
Brinkmeyer L, Drumond RR, Scholz R, Grabocka J, Schmidt-Thieme L (2019) Chameleon: Learning model initializations across tasks with different schemas. arXiv preprint arXiv:190913576
2019
Closest in time.