Fetching the paper…
Reading the bibliography…
We learn recurrent neural network optimizers trained on simple synthetic functions by gradient descent.
Reminiscence and rote learning
L. B. Ward · 1937
Earlier work this paper cites.
The formation of learning sets
H. F. Harlow · 1949
Earlier work this paper cites.
The Bayesian approach to global optimization
J. Močkus · 1982
Earlier work this paper cites.
Evolutionary Principles in Self-Referential Learning. On Learning how to Learn: The Meta-Meta-Meta…-Hook
J. Schmidhuber · 1987
Earlier work this paper cites.
A layered network model of associative learning: learning to learn and configuration
E. J. Kehoe · 1988
Earlier work this paper cites.
On the optimization of a synaptic learning rule
Y. Bengio, S. Bengio, J. Cloutier, and J. Gecsei · 1992
Earlier work this paper cites.
Meta-neural networks that learn by learning
D. K. Naik and R. Mammone · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Learning to learn
S. Thrun and L. Pratt · 1998
Earlier work this paper cites.
Learning to learn using gradient descent
S. Hochreiter, A. S. Younger, and P. R. Conwell · 2001
Earlier work this paper cites.
Gaussian Processes for Machine Learning
C. E. Rasmussen and C. K. I. Williams · 2006
Earlier work this paper cites.
Core knowledge
E. S. Spelke and K. D. Kinzler · 2007
Cited alongside, same era.
E. Brochu, V. M. Cora, and N. de Freitas · 2009
Cited alongside, same era.
Pure exploration in multi-armed bandits problems
S. Bubeck, R. Munos, and G. Stoltz · 2009
Cited alongside, same era.
New inference strategies for solving Markov decision processes using reversible jump MCMC
M. W. Hoffman, H. Kueck, N. de Freitas, and A. Doucet · 2009
Cited alongside, same era.
Controlled experiments on the web: survey and practical guide
R. Kohavi, R. Longbotham, D. Sommerfield, and R. M. Henne · 2009
Cited alongside, same era.
A modern Bayesian look at the multi-armed bandit
S. L. Scott · 2010
Parallelizing exploration-exploitation tradeoffs with Gaussian process bandit optimization
T. Desautels, A. Krause, and J. Burdick · 2014
Later among the works it cites.
Input warping for Bayesian optimization of non-stationary functions
J. Snoek, K. Swersky, R. S. Zemel, and R. P. Adams · 2014
Later among the works it cites.
Bayesian multi-scale optimistic optimization
Z. Wang, B. Shakibi, L. Jin, and N. de Freitas · 2014
Later among the works it cites.
Learning to learn by gradient descent by gradient descent
M. Andrychowicz, M. Denil, S. Gomez, M. W. Hoffman, D. Pfau, T. Schaul, B. Shillingford, and N. de Freitas · 2016
Closest in time.
Hybrid computing using a neural network with dynamic external memory
A. Graves, G. Wayne, M. Reynolds, T. Harley, I. Danihelka, A. Grabska-BarwiÅska, S. G. Colmenarejo, E. Grefenstette, T. Ramalho, J. Agapiou, A. A. P. Badia, K. M. Hermann, Y. Zwols, G. Ostrovski, A. Cain, H. King, C. Summerfield, P. Blunsom, K. Kavukcuoglu, and D. Hassabis · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Gaussian process optimization in the bandit setting: No regret and experimental design
N. Srinivas, A. Krause, S. M. Kakade, and M. Seeger · 2010
Cited alongside, same era.
Algorithms for hyper-parameter optimization
J. S. Bergstra, R. Bardenet, Y. Bengio, and B. Kégl · 2011
Cited alongside, same era.
Best arm identification: A unified approach to fixed budget and fixed confidence
V. Gabillon, M. Ghavamzadeh, and A. Lazaric · 2012
Cited alongside, same era.
Practical Bayesian optimization of machine learning algorithms
J. Snoek, H. Larochelle, and R. P. Adams · 2012
Cited alongside, same era.
Towards an empirical foundation for assessing bayesian optimization of hyperparameters
K. Eggensperger, M. Feurer, F. Hutter, J. Bergstra, J. Snoek, H. Hoos, and K. Leyton-Brown · 2013
Cited alongside, same era.
Sequential model-based optimization for general algorithm configuration
F. Hutter, H. H. Hoos, and K. Leyton-Brown
Cited in the paper.
Meta-learning with memory-augmented neural networks
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. Lillicrap · 2016
Closest in time.
Taking the human out of the loop: A review of Bayesian optimization
B. Shahriari, K. Swersky, Z. Wang, R. P. Adams, and N. de Freitas · 2016
Closest in time.
Learning to reinforcement learn
J. X. Wang, Z. Kurth-Nelson, D. Tirumala, H. Soyer, J. Z. Leibo, R. Munos, C. Blundell, D. Kumaran, and M. Botvinick · 2016
Closest in time.
Learning to optimize
S. Li and J. Malik · 2017
Closest in time.
Optimization as a model for few-shot learning
S. Ravi and H. Larochelle · 2017
Closest in time.
Neural architecture search with reinforcement learning
B. Zoph and Q. V. Le · 2017
Closest in time.