Fetching the paper…
Reading the bibliography…
We propose an algorithm for meta-learning that is model-agnostic, in the sense that it is compatible with any model trained with gradient descent and applicable to a variety of different learning problems, including classification, regression, and reinforcement learning.
Evolutionary principles in self-referential learning
Schmidhuber, Jurgen · 1987
Earlier work this paper cites.
Learning a synaptic learning rule
Bengio, Yoshua, Bengio, Samy, and Cloutier, Jocelyn · 1990
Earlier work this paper cites.
On the optimization of a synaptic learning rule
Bengio, Samy, Bengio, Yoshua, Cloutier, Jocelyn, and Gecsei, Jan · 1992
Earlier work this paper cites.
Meta-neural networks that learn by learning
Naik, Devang K and Mammone, RJ · 1992
Earlier work this paper cites.
Learning to control fast-weight memories: An alternative to dynamic recurrent networks
Schmidhuber, Jürgen · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, Ronald J · 1992
Earlier work this paper cites.
Learning to learn
Thrun, Sebastian and Pratt, Lorien · 1998
Earlier work this paper cites.
Fast learning for problem classes using knowledge based network initialization
Husken, Michael and Goerick, Christian · 2000
Earlier work this paper cites.
Learning to learn using gradient descent
Hochreiter, Sepp, Younger, A Steven, and Conwell, Peter R · 2001
Earlier work this paper cites.
One shot learning of simple visual concepts
Lake, Brenden M, Salakhutdinov, Ruslan, Gross, Jason, and Tenenbaum, Joshua B · 2011
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Todorov, Emanuel, Erez, Tom, and Tassa, Yuval · 2012
Earlier work this paper cites.
Decaf: A deep convolutional activation feature for generic visual recognition
Donahue, Jeff, Jia, Yangqing, Vinyals, Oriol, Hoffman, Judy, Zhang, Ning, Tzeng, Eric, and Darrell, Trevor · 2014
Earlier work this paper cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Saxe, Andrew, McClelland, James, and Ganguli, Surya · 2014
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Goodfellow, Ian J, Shlens, Jonathon, and Szegedy, Christian · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, Sergey and Szegedy, Christian · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, Diederik and Ba, Jimmy · 2015
Cited alongside, same era.
Siamese neural networks for one-shot image recognition
Koch, Gregory · 2015
Cited alongside, same era.
Gradient-based hyperparameter optimization through reversible learning
Maclaurin, Dougal, Duvenaud, David, and Adams, Ryan · 2015
Cited alongside, same era.
Online representation learning in recurrent neural language models
Rei, Marek · 2015
Weight normalization: A simple reparameterization to accelerate training of deep neural networks
Salimans, Tim and Kingma, Diederik P · 2016
Later among the works it cites.
Meta-learning with memory-augmented neural networks
Santoro, Adam, Bartunov, Sergey, Botvinick, Matthew, Wierstra, Daan, and Lillicrap, Timothy · 2016
Later among the works it cites.
Matching networks for one shot learning
Vinyals, Oriol, Blundell, Charles, Lillicrap, Tim, Wierstra, Daan, et al · 2016
Later among the works it cites.
Learning to reinforcement learn
Wang, Jane X, Kurth-Nelson, Zeb, Tirumala, Dhruva, Soyer, Hubert, Leibo, Joel Z, Munos, Remi, Blundell, Charles, Kumaran, Dharshan, and Botvinick, Matt · 2016
Later among the works it cites.
Towards a neural statistician
Edwards, Harrison and Storkey, Amos · 2017
Closest in time.
Hypernetworks
Ha, David, Dai, Andrew, and Le, Quoc V · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Trust region policy optimization
Schulman, John, Levine, Sergey, Abbeel, Pieter, Jordan, Michael I, and Moritz, Philipp · 2015
Cited alongside, same era.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Abadi, Martín, Agarwal, Ashish, Barham, Paul, Brevdo, Eugene, Chen, Zhifeng, Citro, Craig, Corrado, Greg S, Davis, Andy, Dean, Jeffrey, Devin, Matthieu, et al · 2016
Cited alongside, same era.
Learning to learn by gradient descent by gradient descent
Andrychowicz, Marcin, Denil, Misha, Gomez, Sergio, Hoffman, Matthew W, Pfau, David, Schaul, Tom, and de Freitas, Nando · 2016
Cited alongside, same era.
Overcoming catastrophic forgetting in neural networks
Kirkpatrick, James, Pascanu, Razvan, Rabinowitz, Neil, Veness, Joel, Desjardins, Guillaume, Rusu, Andrei A, Milan, Kieran, Quan, John, Ramalho, Tiago, Grabska-Barwinska, Agnieszka, et al · 2016
Cited alongside, same era.
Data-dependent initializations of convolutional neural networks
Krähenbühl, Philipp, Doersch, Carl, Donahue, Jeff, and Darrell, Trevor · 2016
Cited alongside, same era.
Actor-mimic: Deep multitask and transfer reinforcement learning
Parisotto, Emilio, Ba, Jimmy Lei, and Salakhutdinov, Ruslan · 2016
Cited alongside, same era.
Closest in time.
Learning to remember rare events
Kaiser, Lukasz, Nachum, Ofir, Roy, Aurko, and Bengio, Samy · 2017
Closest in time.
Learning to optimize
Li, Ke and Malik, Jitendra · 2017
Closest in time.
Meta networks
Munkhdalai, Tsendsuren and Yu, Hong · 2017
Closest in time.
Optimization as a model for few-shot learning
Ravi, Sachin and Larochelle, Hugo · 2017
Closest in time.
Attentive recurrent comparators
Shyam, Pranav, Gupta, Shubham, and Dukkipati, Ambedkar · 2017
Closest in time.
Prototypical networks for few-shot learning
Snell, Jake, Swersky, Kevin, and Zemel, Richard S · 2017
Closest in time.