Fetching the paper…
Reading the bibliography…
In modern supervised learning, there are a large number of tasks, but many of them are associated with only a small amount of labeled data.
Probability inequalities for sums of bounded random variables
Hoeffding, W · 1963
Earlier work this paper cites.
Estimating true-score distributions in psychological testing (an empirical bayes estimation problem)
Lord, F. M · 1969
Earlier work this paper cites.
Slink: an optimally efficient algorithm for the single-link cluster method
Sibson, R · 1973
Earlier work this paper cites.
Evolutionary principles in self-referential learning, or on learning how to learn: the meta-meta-… hook
Schmidhuber, J · 1987
Earlier work this paper cites.
A model of inductive bias learning
Baxter, J · 2000
Earlier work this paper cites.
A spectral algorithm for learning mixture models
Vempala, S. and Wang, G · 2004
Earlier work this paper cites.
A framework for learning predictive structures from multiple tasks and unlabeled data
Ando, R. K. and Zhang, T · 2005
Earlier work this paper cites.
Supervised dimensionality reduction using mixture models
Orlitsky, A · 2005
Earlier work this paper cites.
Uncovering shared structures in multiclass classification
Amit, Y., Fink, M., Srebro, N., and Ullman, S · 2007
Earlier work this paper cites.
Convex multi-task feature learning
Argyriou, A., Evgeniou, T., and Pontil, M · 2008
Earlier work this paper cites.
Closed-form supervised dimensionality reduction with generalized linear models
Rish, I., Grabarnik, G., Cecchi, G., Pereira, F., and Gordon, G. J · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Bayes and empirical Bayes methods for data analysis
Carlin, B. P. and Louis, T. A · 2010
Earlier work this paper cites.
Large-scale inference: empirical Bayes methods for estimation, testing, and prediction , volume 1
Efron, B · 2012
Earlier work this paper cites.
Large-scale image classification with trace-norm regularization
Harchaoui, Z., Douze, M., Paulin, M., Dudik, M., and Malick, J · 2012
Earlier work this paper cites.
Random design analysis of ridge regression
Hsu, D., Kakade, S. M., and Zhang, T · 2012
Earlier work this paper cites.
Learning to learn
Thrun, S. and Pratt, L · 2012
Earlier work this paper cites.
Spectral experts for estimating mixtures of linear regressions
Chaganty, A. T. and Liang, P · 2013
Earlier work this paper cites.
Excess risk bounds for multitask learning with trace norm regularization
Pontil, M. and Maurer, A · 2013
Cited alongside, same era.
Findings of the 2014 workshop on statistical machine translation
Bojar, O., Buck, C., Federmann, C., Haddow, B., Koehn, P., Leveling, J., Monz, C., Pecina, P., Post, M., Saint-Amand, H., et al · 2014
Cited alongside, same era.
Efficient representations for lifelong learning and autoencoding
Balcan, M.-F., Blum, A., and Vempala, S · 2015
Cited alongside, same era.
Siamese neural networks for one-shot image recognition
Koch, G., Zemel, R., and Salakhutdinov, R · 2015
Cited alongside, same era.
An introduction to matrix concentration inequalities
Tropp, J. A. et al · 2015
Cited alongside, same era.
Lazysvd: even faster svd decomposition yet without agonizing pain
Mixture models, robustness, and sum of squares proofs
Hopkins, S. B. and Li, J · 2018
Later among the works it cites.
Bayesian model-agnostic meta-learning
Kim, T., Yoon, J., Dia, O., Kim, S., Bengio, Y., and Ahn, S · 2018
Later among the works it cites.
Estimating learnability in the sublinear data regime
Kong, W. and Valiant, G · 2018
Later among the works it cites.
Robust moment estimation and improved clustering via sum of squares
Kothari, P. K., Steinhardt, J., and Steurer, D · 2018
Later among the works it cites.
Learning mixtures of linear regressions with nearly optimal complexity
Li, Y. and Liang, Y · 2018
Later among the works it cites.
Tadam: Task dependent adaptive metric for improved few-shot learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Allen-Zhu, Z. and Li, Y · 2016
Cited alongside, same era.
Optimization as a model for few-shot learning
Ravi, S. and Larochelle, H · 2016
Cited alongside, same era.
Provable tensor methods for learning mixtures of generalized linear models
Sedghi, H., Janzamin, M., and Anandkumar, A · 2016
Cited alongside, same era.
Yi, X., Caramanis, C., and Sanghavi, S · 2016
Cited alongside, same era.
Mixed linear regression with multiple components
Zhong, K., Jain, P., and Dhillon, I. S · 2016
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S · 2017
Cited alongside, same era.
Meta-sgd: Learning to learn quickly for few-shot learning
Li, Z., Zhou, F., Chen, F., and Li, H · 2017
Cited alongside, same era.
Oreshkin, B., López, P. R., and Lacoste, A · 2018
Later among the works it cites.
Meta-learning with latent embedding optimization
Rusu, A. A., Rao, D., Sygnowski, J., Vinyals, O., Pascanu, R., Osindero, S., and Hadsell, R · 2018
Later among the works it cites.
High-dimensional probability: An introduction with applications in data science , volume 47
Vershynin, R · 2018
Later among the works it cites.
Deep meta-learning: Learning to learn in the concept space
Zhou, F., Wu, B., and Li, Z · 2018
Later among the works it cites.
Meta-learning with differentiable closed-form solvers
Bertinetto, L., Henriques, J. F., Torr, P. H., and Vedaldi, A · 2019
Later among the works it cites.
Generalize across tasks: Efficient algorithms for linear representation learning
Bullins, B., Hazan, E., Kalai, A., and Livni, R · 2019
Later among the works it cites.
Sublinear optimal policy value estimation in contextual bandits
Kong, W., Valiant, G., and Brunskill, E · 2019
Later among the works it cites.
Meta-learning with implicit gradients
Rajeswaran, A., Finn, C., Kakade, S. M., and Levine, S · 2019
Later among the works it cites.
Meta-dataset: A dataset of datasets for learning to learn from few examples
Triantafillou, E., Zhu, T., Dumoulin, V., Lamblin, P., Xu, K., Goroshin, R., Gelada, C., Swersky, K., Manzagol, P.-A., and Larochelle, H · 2019
Later among the works it cites.
Maximum likelihood estimation for learning populations of parameters
Vinayak, R. K., Kong, W., Valiant, G., and Kakade, S. M · 2019
Later among the works it cites.
Fast context adaptation via meta-learning
Zintgraf, L., Shiarli, K., Kurin, V., Hofmann, K., and Whiteson, S · 2019
Later among the works it cites.
Learning mixtures of linear regressions in subexponential time via Fourier moments
Chen, S., Li, J., and Song, Z · 2020
Closest in time.