Fetching the paper…
Reading the bibliography…
We present a self-contained system for constructing natural language models for use in text compression.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, N., Hinton, G. E., Krizhevsky, A., Sutskever, I., and Salakhutdinov, R · 1958
Earlier work this paper cites.
Three approaches to the quantitative definition ofinformation’
Kolmogorov, A. N · 1965
Earlier work this paper cites.
Project gutenberg
Hart, M · 1971
Earlier work this paper cites.
Source coding algorithms for fast data compression
Pasco, R. C · 1976
Earlier work this paper cites.
Data compression using adaptive coding and partial string matching
Cleary, J., and Witten, I · 1984
Earlier work this paper cites.
Arithmetic coding for data compression
Witten, I. H., Neal, R. M., and Cleary, J. G · 1987
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
Werbos, P. J · 1990
Earlier work this paper cites.
The problem of learning long-term dependencies in recurrent networks
Bengio, Y., Frasconi, P., and Simard, P · 1993
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Y., Simard, P., and Frasconi, P · 1994
Earlier work this paper cites.
On the dynamics of small continuous-time recurrent neural networks
Beer, R. D · 1995
Earlier work this paper cites.
Long short-term memory
Hochreiter, S., and Schmidhuber, J · 1997
Cited alongside, same era.
A guide to recurrent neural networks and backpropagation
Boden, M · 2002
Cited alongside, same era.
The paq1 data compression program
Mahoney, M. V · 2002
Cited alongside, same era.
Theano: A cpu and gpu math compiler in python
Bergstra, J., Breuleux, O., Bastien, F., Lamblin, P., Pascanu, R., Desjardins, G., Turian, J., Warde-Farley, D., and Bengio, Y · 2010
Cited alongside, same era.
Understanding the difficulty of training deep feedforward neural networks
Glorot, X., and Bengio, Y · 2010
Cited alongside, same era.
Data Compression Explained
Mahoney, M · 2010
Cited alongside, same era.
Lstm neural networks for language modeling
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Later among the works it cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Chung, J., Gulcehre, C., Cho, K., and Bengio, Y · 2014
Later among the works it cites.
Long short-term memory recurrent neural network architectures for large scale acoustic modeling
Sak, H., Senior, A. W., and Beaufays, F · 2014
Later among the works it cites.
Gated feedback recurrent neural networks
Chung, J., Gülçehre, C., Cho, K., and Bengio, Y · 2015
Later among the works it cites.
Long-term recurrent convolutional networks for visual recognition and description
Donahue, J., Anne Hendricks, L., Guadarrama, S., Rohrbach, M., Venugopalan, S., Saenko, K., and Darrell, T · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sundermeyer, M., Schlüter, R., and Ney, H · 2012
Cited alongside, same era.
Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude
Tieleman, T., and Hinton, G · 2012
Cited alongside, same era.
How to construct deep recurrent neural networks
Pascanu, R., Gulcehre, C., Cho, K., and Bengio, Y · 2013
Cited alongside, same era.
On the difficulty of training recurrent neural networks
Pascanu, R., Mikolov, T., and Bengio, Y · 2013
Cited alongside, same era.
Keras: Deep learning library for theano and tensorflow
Cited in the paper.
Fast text compression with neural networks
Mahoney, M. V
Cited in the paper.
Later among the works it cites.
A primer on neural network models for natural language processing
Goldberg, Y · 2015
Later among the works it cites.
Visualizing and understanding recurrent networks
Karpathy, A., Johnson, J., and Fei-Fei, L · 2015
Later among the works it cites.
Convolutional, long short-term memory, fully connected deep neural networks
Sainath, T. N., Vinyals, O., Senior, A., and Sak, H · 2015
Later among the works it cites.
Going deeper with convolutions
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V., and Rabinovich, A · 2015
Later among the works it cites.
Globally normalized transition-based neural networks
Andor, D., Alberti, C., Weiss, D., Severyn, A., Presta, A., Ganchev, K., Petrov, S., and Collins, M · 2016
Closest in time.