Fetching the paper…
Reading the bibliography…
Current deep learning based text classification methods are limited by their ability to achieve fast learning and generalization when the data is scarce.
Clues to suicide
Edwin S Shneidman and Norman L Farberow · 1956
Earlier work this paper cites.
The effect of background knowledge on the reading comprehension of second language learners
Martin G Levine and George J Haus · 1985
Earlier work this paper cites.
Evolutionary principles in self-referential learning. on learning now to learn: The meta-meta-meta…-hook
Jurgen Schmidhuber · 1987
Earlier work this paper cites.
On the optimization of a synaptic learning rule
Samy Bengio, Yoshua Bengio, Jocelyn Cloutier, and Jan Gecsei · 1992
Earlier work this paper cites.
Explanation-based neural network learning for robot control
Tom M Mitchell and Sebastian B Thrun · 1993
Earlier work this paper cites.
Bidirectional recurrent neural networks
Mike Schuster and Kuldip K Paliwal · 1997
Earlier work this paper cites.
Lifelong learning algorithms
Sebastian Thrun · 1998
Earlier work this paper cites.
Learning to learn using gradient descent
Sepp Hochreiter, A Steven Younger, and Peter R Conwell · 2001
Earlier work this paper cites.
A perspective view and survey of meta-learning
Ricardo Vilalta and Youssef Drissi · 2002
Earlier work this paper cites.
Rcv1: A new benchmark collection for text categorization research
David D Lewis, Yiming Yang, Tony G Rose, and Fan Li · 2004
Earlier work this paper cites.
Evolving modular fast-weight networks for control
Faustino Gomez and Jürgen Schmidhuber · 2005
Earlier work this paper cites.
One-shot learning of object categories
Fei-Fei Li, Rob Fergus, and Pietro Perona · 2006
Earlier work this paper cites.
One shot learning of simple visual concepts
Brenden Lake, Ruslan Salakhutdinov, Jason Gross, and Joshua Tenenbaum · 2011
Earlier work this paper cites.
Active learning for clinical text classification: is it better than random sampling?
Rosa L Figueroa, Qing Zeng-Treitler, Long H Ngo, Sergey Goryachev, and Eduardo P Wiechmann · 2012
Earlier work this paper cites.
Effect of small sample size on text categorization with support vector machines
Pawel Matykiewicz and John Pestian · 2012
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Andrew L Maas, Awni Y Hannun, and Andrew Y Ng · 2013
Cited alongside, same era.
Regularization of neural networks using dropconnect
Li Wan, Matthew Zeiler, Sixin Zhang, Yann Le Cun, and Rob Fergus · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
Alex Graves, Greg Wayne, and Ivo Danihelka · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Cited alongside, same era.
A decomposable attention model for natural language inference
Ankur P Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit · 2016
Later among the works it cites.
Optimization as a model for few-shot learning
Sachin Ravi and Hugo Larochelle · 2016
Later among the works it cites.
Meta-learning with memory-augmented neural networks
Adam Santoro, Sergey Bartunov, Matthew Botvinick, Daan Wierstra, and Timothy Lillicrap · 2016
Later among the works it cites.
Wavenet: A generative model for raw audio
Aaron Van Den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Matching networks for one shot learning
Oriol Vinyals, Charles Blundell, Tim Lillicrap, Daan Wierstra, et al · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning · 2014
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Cited alongside, same era.
A review on multi-label learning algorithms
Min-Ling Zhang and Zhi-Hua Zhou · 2014
Cited alongside, same era.
Teaching machines to read and comprehend
Karl Moritz Hermann, Tomas Kocisky, Edward Grefenstette, Lasse Espeholt, Will Kay, Mustafa Suleyman, and Phil Blunsom · 2015
Cited alongside, same era.
Siamese neural networks for one-shot image recognition
Gregory Koch, Richard Zemel, and Ruslan Salakhutdinov · 2015
Cited alongside, same era.
Human-level concept learning through probabilistic program induction
Brenden M Lake, Ruslan Salakhutdinov, and Joshua B Tenenbaum · 2015
Cited alongside, same era.
Feed-forward networks with attention can solve some long-term memory problems
Colin Raffel and Daniel PW Ellis · 2015
Cited alongside, same era.
Minmin Chen · 2017
Later among the works it cites.
Chelsea Finn and Sergey Levine · 2017
Later among the works it cites.
Learning to optimize neural nets
Ke Li and Jitendra Malik · 2017
Later among the works it cites.
Meta-sgd: Learning to learn quickly for few shot learning
Zhenguo Li, Fengwei Zhou, Fei Chen, and Hang Li · 2017
Later among the works it cites.
A structured self-attentive sentence embedding
Zhouhan Lin, Minwei Feng, Cicero Nogueira dos Santos, Mo Yu, Bing Xiang, Bowen Zhou, and Yoshua Bengio · 2017
Later among the works it cites.
Meta-learning with temporal convolutions
Nikhil Mishra, Mostafa Rohaninejad, Xi Chen, and Pieter Abbeel · 2017
Later among the works it cites.
Meta networks
Tsendsuren Munkhdalai and Hong Yu · 2017
Later among the works it cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Later among the works it cites.
An empirical evaluation of generic convolutional and recurrent networks for sequence modeling
Shaojie Bai, J Zico Kolter, and Vladlen Koltun · 2018
Closest in time.
Bi-directional block self-attention for fast and memory-efficient sequence modeling
Tao Shen, Tianyi Zhou, Guodong Long, Jing Jiang, and Chengqi Zhang · 2018
Closest in time.