Torchmeta: A Meta-Learning library for PyTorch, 2019
Original
T. Deleu, T. Würfl, M. Samiei, J. P. Cohen, and Y. Bengio · 1909
Earlier work this paper cites.
Evolutionary principles in self-referential learning
J. Schmidhuber · 1987
Earlier work this paper cites.
Building symmetries into feedforward networks
J. Shawe-Taylor · 1989
Earlier work this paper cites.
On the optimization of a synaptic learning rule
S. Bengio, Y. Bengio, J. Cloutier, and J. Gecsei · 1992
Earlier work this paper cites.
Learning to control fast-weight memories: An alternative to dynamic recurrent networks
J. Schmidhuber · 1992
Earlier work this paper cites.
Face recognition from one example view
D. Beymer and T. Poggio · 1995
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Incorporating prior information in machine learning by creating virtual examples
P. Niyogi, F. Girosi, and T. Poggio · 1998
Earlier work this paper cites.
Learning to learn using gradient descent
S. Hochreiter, A. S. Younger, and P. R. Conwell · 2001
Earlier work this paper cites.
Evolving neural networks through augmenting topologies
K. O. Stanley and R. Miikkulainen · 2002
Earlier work this paper cites.
Abstract algebra , volume 3
D. S. Dummit and R. M. Foote · 2004
Earlier work this paper cites.
Tensor decompositions and applications
T. G. Kolda and B. W. Bader · 2009
Earlier work this paper cites.
A hypercube-based encoding for evolving large-scale neural networks
K. O. Stanley, D. B. D’Ambrosio, and J. Gauci · 2009
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Learning to learn
S. Thrun and L. Pratt · 2012
Earlier work this paper cites.
Deep symmetry networks
R. Gens and P. M. Domingos · 2014
Earlier work this paper cites.
Towards end-to-end speech recognition with recurrent neural networks
A. Graves and N. Jaitly · 2014
Earlier work this paper cites.
Deep speech: Scaling up end-to-end speech recognition
Original
A. Hannun, C. Case, J. Casper, B. Catanzaro, G. Diamos, E. Elsen, R. Prenger, S. Satheesh, S. Sengupta, A. Coates, et al · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Original
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.