Fetching the paper…
Reading the bibliography…
This paper introduces Grid Long Short-Term Memory, a network of LSTM cells arranged in a multidimensional grid that can be applied to vectors, sequences or higher dimensional data such as images.
Perceptrons: An Introduction to Computational Geometry
Marvin Minsky, Seymour Papert · 1972
Earlier work this paper cites.
Untersuchungen zu dynamischen neuronalen Netzen. Diploma thesis, Institut für Informatik, Lehrstuhl Prof. Brauer, Technische Universität München, 1991
Hochreiter, S · 1991
Earlier work this paper cites.
Long Short-Term Memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Solving the n-bit parity problem using neural networks
Hohil, Myron E., Liu, Derong, and Smith, Stanley H · 1999
Earlier work this paper cites.
Gradient flow in recurrent nets: the difficulty of learning long-term dependencies
Hochreiter, S., Bengio, Y., Frasconi, P., and Schmidhuber, J · 2001
Earlier work this paper cites.
Best practices for convolutional neural networks applied to visual document analysis
Simard, Patrice Y., Steinkraus, David, and Platt, John C · 2003
Earlier work this paper cites.
K-separability
Duch, Wlodzislaw · 2006
Earlier work this paper cites.
Multi-dimensional recurrent neural networks
Graves, A., Fernández, S., and Schmidhuber, J · 2007
Earlier work this paper cites.
Offline handwriting recognition with multidimensional recurrent neural networks
Graves, A. and Schmidhuber, J · 2008
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
Duchi, John, Hazan, Elad, and Singer, Yoram · 2010
Earlier work this paper cites.
cdec: A decoder, alignment, and learning framework for finite-state and context-free translation models
Dyer, Chris, Lopez, Adam, Ganitkevitch, Juri, Weese, Johnathan, Ture, Ferhan, Blunsom, Phil, Setiawan, Hendra, Eidelman, Vladimir, and Resnik, Philip · 2010
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, Vinod and Hinton, Geoffrey E · 2010
Cited alongside, same era.
Generating text with recurrent neural networks
Sutskever, I., Martens, J., and Hinton, G · 2011
Cited alongside, same era.
Multi-column deep neural networks for image classification
Ciresan, Dan Claudiu, Meier, Ueli, and Schmidhuber, Jürgen · 2012
Cited alongside, same era.
Supervised sequence labelling with recurrent neural networks , volume 385
Graves, A · 2012
Cited alongside, same era.
The human knowledge compression context, 2012
Hutter, Marcus · 2012
Cited alongside, same era.
Maxout networks
Goodfellow, Ian J., Warde-Farley, David, Mirza, Mehdi, Courville, Aaron C., and Bengio, Yoshua · 2013
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Cho, Kyunghyun, van Merrienboer, Bart, Gülçehre, Çaglar, Bougares, Fethi, Schwenk, Holger, and Bengio, Yoshua · 2014
Later among the works it cites.
Adam: A method for stochastic optimization
Kingma, Diederik P. and Ba, Jimmy · 2014
Later among the works it cites.
Unifying visual-semantic embeddings with multimodal neural language models
Kiros, Ryan, Salakhutdinov, Ruslan, and Zemel, Richard S · 2014
Later among the works it cites.
Mairal, Julien, Koniusz, Piotr, Harchaoui, Zaïd, and Schmid, Cordelia · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Speech recognition with deep recurrent neural networks
Graves, A., Mohamed, A., and Hinton, G · 2013
Cited alongside, same era.
Generating sequences with recurrent neural networks
Graves, Alex · 2013
Cited alongside, same era.
Recurrent continuous translation models
Kalchbrenner, Nal and Blunsom, Phil · 2013
Cited alongside, same era.
Lin, Min, Chen, Qiang, and Yan, Shuicheng · 2013
Cited alongside, same era.
Regularization of neural networks using dropconnect
Wan, Li, Zeiler, Matthew D., Zhang, Sixin, LeCun, Yann, and Fergus, Rob · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2014
Cited alongside, same era.
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc VV · 2014
Later among the works it cites.
Show and tell: A neural image caption generator
Vinyals, Oriol, Toshev, Alexander, Bengio, Samy, and Erhan, Dumitru · 2014
Later among the works it cites.
Zaremba, Wojciech and Sutskever, Ilya · 2014
Later among the works it cites.
Gated feedback recurrent neural networks
Chung, Junyoung, Gülçehre, Çaglar, Cho, KyungHyun, and Bengio, Yoshua · 2015
Closest in time.
Deeply-supervised nets
Lee, Chen-Yu, Xie, Saining, Gallagher, Patrick, Zhang, Zhengyou, and Tu, Zhuowen · 2015
Closest in time.
Srivastava, Rupesh Kumar, Greff, Klaus, and Schmidhuber, Jürgen · 2015
Closest in time.
Renet: A recurrent neural network based alternative to convolutional networks
Visin, Francesco, Kastner, Kyle, Cho, Kyunghyun, Matteucci, Matteo, Courville, Aaron C., and Bengio, Yoshua · 2015
Closest in time.
Yao, Kaisheng, Cohn, Trevor, Vylomova, Katerina, Duh, Kevin, and Dyer, Chris · 2015
Closest in time.