Fetching the paper…
Reading the bibliography…
We present two approaches that use unlabeled data to improve sequence learning with recurrent networks.
Beyond regression: New tools for prediction and analysis in the behavioral sciences
P. J. Werbos · 1974
Earlier work this paper cites.
Learning representations by back-propagating errors
D. Rumelhart, G. E. Hinton, and R. J. Williams · 1986
Earlier work this paper cites.
Newsweeder: Learning to filter netnews
K. Lang · 1995
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Learning to forget: Continual prediction with LSTM
F. A. Gers, J. Schmidhuber, and F. Cummins · 2000
Earlier work this paper cites.
Gradient flow in recurrent nets: the difficulty of learning long-term dependencies
S. Hochreiter, Y. Bengio, P. Frasconi, and J. Schmidhuber · 2001
Earlier work this paper cites.
A framework for learning predictive structures from multiple tasks and unlabeled data
R. K. Ando and T. Zhang · 2005
Earlier work this paper cites.
Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales
B. Pang and L. Lee · 2005
Earlier work this paper cites.
Convolutional deep belief networks on CIFAR-10
A. Krizhevsky · 2010
Earlier work this paper cites.
Recurrent neural network based language model
T. Mikolov, M. Karafiát, L. Burget, J. Cernockỳ, and S. Khudanpur · 2010
Earlier work this paper cites.
Learning word vectors for sentiment analysis
A. L. Maas, R. E. Daly, P. T. Pham, D. Huang, A. Y. Ng, and C. Potts · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Learning algorithms for the classification restricted boltzmann machine
H. Larochelle, M. Mandel, R. Pascanu, and Y. Bengio · 2012
Earlier work this paper cites.
Semantic compositionality through recursive matrix-vector spaces
R. Socher, B. Huval, C. D. Manning, and A. Y. Ng · 2012
Cited alongside, same era.
Baselines and bigrams: Simple, good sentiment and topic classification
S. I. Wang and C. D. Manning · 2012
Cited alongside, same era.
Stochastic ratio matching of RBMs for sparse high-dimensional inputs
Y. Dauphin and Y. Bengio · 2013
Cited alongside, same era.
Hidden factors and hidden topics: understanding rating dimensions with review text
J. McAuley and J. Leskovec · 2013
Cited alongside, same era.
Recursive deep models for semantic compositionality over a sentiment treebank
R. Socher, A. Perelygin, J. Y. Wu, J. Chuang, C. D. Manning, A. Y. Ng, and C. Potts · 2013
Cited alongside, same era.
End-to-end continuous speech recognition using attention-based recurrent nn: First results
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2014
Later among the works it cites.
Recurrent neural network regularization
W. Zaremba, I. Sutskever, and O. Vinyals · 2014
Later among the works it cites.
Datasets for single-label text categorization
A. Cardoso-Cachopo · 2015
Closest in time.
William Chan, Navdeep Jaitly, Quoc V Le, and Oriol Vinyals · 2015
Closest in time.
LSTM: A search space odyssey
K. Greff, R. K. Srivastava, J. Koutník, B. R. Steunebrink, and J. Schmidhuber · 2015
Closest in time.
Skip-thought vectors
R. Kiros, Y. Zhu, R. Salakhutdinov, R. S. Zemel, A. Torralba, R. Urtasun, and S. Fidler · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Chorowski, D. Bahdanau, K. Cho, and Y. Bengio · 2014
Cited alongside, same era.
On using very large target vocabulary for neural machine translation
S. Jean, K. Cho, R. Memisevic, and Y. Bengio · 2014
Cited alongside, same era.
Effective use of word order for text categorization with convolutional neural networks
R. Johnson and T. Zhang · 2014
Cited alongside, same era.
Convolutional neural networks for sentence classification
Y. Kim · 2014
Cited alongside, same era.
Distributed representations of sentences and documents
Q. V. Le and T. Mikolov · 2014
Cited alongside, same era.
DBpedia – a large-scale, multilingual knowledge base extracted from wikipedia
J. Lehmann, R. Isele, M. Jakob, A. Jentzsch, D. Kontokostas, P. N. Mendes, S. Hellmann, M. Morsey, P. van Kleef, S. Auer, et al · 2014
Cited alongside, same era.
Addressing the rare word problem in neural machine translation
T. Luong, I. Sutskever, Q. V. Le, O. Vinyals, and W. Zaremba · 2014
Cited alongside, same era.
Closest in time.
Beyond short snippets: Deep networks for video classification
J. Y. H. Ng, M. J. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici · 2015
Closest in time.
Neural responding machine for short-text conversation
L. Shang, Z. Lu, and H. Li · 2015
Closest in time.
Unsupervised learning of video representations using LSTMs
N. Srivastava, E. Mansimov, and R. Salakhutdinov · 2015
Closest in time.
Grammar as a foreign language
O. Vinyals, L. Kaiser, T. Koo, S. Petrov, I. Sutskever, and G. Hinton · 2015
Closest in time.
A neural conversational model
O. Vinyals and Q. V. Le · 2015
Closest in time.
Character-level convolutional networks for text classification
X. Zhang and Y. LeCun · 2015
Closest in time.