Fetching the paper…
Reading the bibliography…
We study the convergence of a class of gradient-based Model-Agnostic Meta-Learning (MAML) methods and characterize their overall complexity as well as their best achievable accuracy in terms of gradient norm for nonconvex loss functions.
Online meta-learning
Finn, C., Rajeswaran, A., Kakade, S., and Levine, S. (2019) · 1930
Earlier work this paper cites.
Bounds on reciprocal moments with applications and developments in stein estimation and post-stratification
Wooff, D. A. (1985) · 1985
Earlier work this paper cites.
Learning a synaptic learning rule
Bengio, Y., Bengio, S., and Cloutier, J. (1990) · 1990
Earlier work this paper cites.
On the optimization of a synaptic learning rule
Bengio, S., Bengio, Y., Cloutier, J., and Gecsei, J. (1992) · 1992
Earlier work this paper cites.
Learning to learn
Thrun, S. and Pratt, L. (1998) · 1998
Earlier work this paper cites.
A model of inductive bias learning
Baxter, J. (2000) · 2000
Earlier work this paper cites.
Introductory Lectures on Convex Optimization: A Basic Course
Nesterov, Y. (2004) · 2004
Earlier work this paper cites.
Algorithms for hyper-parameter optimization
Bergstra, J. S., Bardenet, R., Bengio, Y., and Kégl, B. (2011) · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
Bergstra, J. and Bengio, Y. (2012) · 2012
Earlier work this paper cites.
Efficient representations for lifelong learning and autoencoding
Balcan, M.-F., Blum, A., and Vempala, S. (2015) · 2015
Earlier work this paper cites.
Learning to learn by gradient descent by gradient descent
Andrychowicz, M., Denil, M., Gómez, S., Hoffman, M. W., Pfau, D., Schaul, T., Shillingford, B., and de Freitas, N. (2016) · 2016
Earlier work this paper cites.
Regret Bounds for Lifelong Learning
Alquier, P., Mai, T. T., and Pontil, M. (2017) · 2017
Cited alongside, same era.
Designing neural network architectures using reinforcement learning
Baker, B., Gupta, O., Naik, N., and Raskar, R. (2017) · 2017
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S. (2017) · 2017
Cited alongside, same era.
Learning to optimize
Li, K. and Malik, J. (2017) · 2017
Cited alongside, same era.
Meta-SGD: Learning to learn quickly for few-shot learning
Li, Z., Zhou, F., Chen, F., and Li, H. (2017) · 2017
Cited alongside, same era.
Optimization as a model for few-shot learning
Ravi, S. and Larochelle, H. (2017) · 2017
Cited alongside, same era.
On first-order meta-learning algorithms
Nichol, A., Achiam, J., and Schulman, J. (2018) · 2018
Later among the works it cites.
Learning transferable architectures for scalable image recognition
Zoph, B., Vasudevan, V., Shlens, J., and Le, Q. V. (2018) · 2018
Later among the works it cites.
How to train your MAML
Antoniou, A., Edwards, H., and Storkey, A. (2019) · 2019
Closest in time.
Alpha MAML: adaptive model-agnostic meta-learning
Behl, H. S., Baydin, A. G., and Torr, P. H. S. (2019) · 2019
Closest in time.
Learning-to-learn stochastic gradient descent with biased regularization
Denevi, G., Ciliberto, C., Grazzi, R., and Pontil, M. (2019) · 2019
Closest in time.
Provable guarantees for gradient-based meta-learning
Khodak, M., Balcan, M.-F., and Talwalkar, A. (2019) · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural architecture search with reinforcement learning
Zoph, B. and Le, Q. V. (2017) · 2017
Cited alongside, same era.
Continuous adaptation via meta-learning in nonstationary and competitive environments
Al-Shedivat, M., Bansal, T., Burda, Y., Sutskever, I., Mordatch, I., and Abbeel, P. (2018) · 2018
Cited alongside, same era.
Learning to learn around a common mean
Denevi, G., Ciliberto, C., Stamos, D., and Pontil, M. (2018) · 2018
Cited alongside, same era.
Bilevel programming for hyperparameter optimization and meta-learning
Franceschi, L., Frasconi, P., Salzo, S., Grazzi, R., and Pontil, M. (2018) · 2018
Cited alongside, same era.
Recasting gradient-based meta-learning as hierarchical bayes
Grant, E., Finn, C., Levine, S., Darrell, T., and Griffiths, T. (2018) · 2018
Cited alongside, same era.
Closest in time.
Learning unsupervised learning rules
Metz, L., Maheswaranathan, N., Cheung, B., and Sohl-Dickstein, J. (2019) · 2019
Closest in time.
Meta-learning with implicit gradients
Rajeswaran, A., Finn, C., Kakade, S. M., and Levine, S. (2019) · 2019
Closest in time.
Meta-Learning
Vanschoren, J. (2019) · 2019
Closest in time.
Fast context adaptation via meta-learning
Zintgraf, L., Shiarli, K., Kurin, V., Hofmann, K., and Whiteson, S. (2019) · 2019
Closest in time.