Fetching the paper…
Reading the bibliography…
We present a neural transducer model with visual attention that learns to generate LaTeX markup of a real-world math formula given its image.
Long short-term memory
Hochreiter, S. and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A., Fernández, S., Gomez, F. J., and Schmidhuber, J. (2006) · 2006
Earlier work this paper cites.
Supervised sequence labelling with recurrent neural networks
Graves, A. (2008) · 2008
Earlier work this paper cites.
Offline handwriting recognition with multidimensional recurrent neural networks
Graves, A. and Schmidhuber, J. (2008) · 2008
Earlier work this paper cites.
Generating sequences with recurrent neural networks
Graves, A. (2013) · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Graves, A., Mohamed, A., and Hinton, G. E. (2013) · 2013
Earlier work this paper cites.
How to construct deep recurrent neural networks
Pascanu, R., Çaglar Gülçehre, Cho, K., and Bengio, Y. (2013) · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
A comparison of sequence-trained deep neural networks and recurrent neural networks optical modeling for handwriting recognition
Bluche, T., Ney, H., and Kermorvant, C. (2014) · 2014
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., van Merrienboer, B., Çaglar Gülçehre, Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J. (2014) · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A. (2014) · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., and Le, Q. V. (2014) · 2014
Cited alongside, same era.
Recurrent neural network regularization
Zaremba, W., Sutskever, I., and Vinyals, O. (2014) · 2014
Cited alongside, same era.
Chan, W., Jaitly, N., Le, Q. V., and Vinyals, O. (2015) · 2015
Show, attend and tell: Neural image caption generation with visual attention
Xu, K., Ba, J., Kiros, J. R., Cho, K., Courville, A. C., Salakhutdinov, R., Zemel, R. S., and Bengio, Y. (2015) · 2015
Later among the works it cites.
Multi-scale context aggregation by dilated convolutions
Yu, F. and Koltun, V. (2015) · 2015
Later among the works it cites.
Joint line segmentation and transcription for end-to-end handwritten paragraph recognition
Bluche, T. (2016) · 2016
Later among the works it cites.
Densecap: Fully convolutional localization networks for dense captioning
Johnson, J., Karpathy, A., and Fei-Fei, L. (2016) · 2016
Later among the works it cites.
Neural machine translation in linear time
Kalchbrenner, N., Espeholt, L., Simonyan, K., van den Oord, A., Graves, A., and Kavukcuoglu, K. (2016) · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Long-term recurrent convolutional networks for visual recognition and description
Donahue, J., Hendricks, L. A., Guadarrama, S., Rohrbach, M., Venugopalan, S., Saenko, K., and Darrell, T. (2015) · 2015
Cited alongside, same era.
Deep visual-semantic alignments for generating image descriptions
Karpathy, A. and fei Li, F. (2015) · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
Luong, M., Pham, H., and Manning, C. D. (2015) · 2015
Cited alongside, same era.
Generative image modeling using spatial lstms
Theis, L. and Bethge, M. (2015) · 2015
Cited alongside, same era.
Show and tell: A neural image caption generator
Vinyals, O., Toshev, A., Bengio, S., and Erhan, D. (2015) · 2015
Cited alongside, same era.
Wavenet: A generative model for raw audio
van den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A. W., and Kavukcuoglu, K. (2016a)
Cited in the paper.
Pixel recurrent neural networks
van den Oord, A., Kalchbrenner, N., and Kavukcuoglu, K. (2016b)
Cited in the paper.
Later among the works it cites.
Areas of attention for image captioning
Pedersoli, M., Lucas, T., Schmid, C., and Verbeek, J. (2016) · 2016
Later among the works it cites.
Image-to-markup generation with coarse-to-fine attention
Deng, Y., Kanervisto, A., Ling, J., and Rush, A. M. (2017) · 2017
Later among the works it cites.
Attention correctness in neural image captioning
Liu, C., Mao, J., Sha, F., and Yuille, A. L. (2017) · 2017
Later among the works it cites.
Salimans, T., Karpathy, A., Chen, X., and Kingma, D. P. (2017) · 2017
Later among the works it cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L., and Polosukhin, I. (2017) · 2017
Later among the works it cites.