Fetching the paper…
Reading the bibliography…
We develop the first approximate inference algorithm for 1-Best (and M-Best) decoding in bidirectional neural sequence models by extending Beam Search (BS) to reason about both forward and backward time dependencies.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
An efficient algorithm for finding the m most probable configurations in probabilistic expert systems
D. Nilsson · 1998
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu · 2002
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
S. Banerjee and A. Lavie · 2005
Earlier work this paper cites.
Parsing natural scenes and natural language with recursive neural networks
R. Socher, C. C. Lin, C. Manning, and A. Y. Ng · 2011
Earlier work this paper cites.
An Efficient Message-Passing Algorithm for the M-Best MAP Problem
D. Batra · 2012
Earlier work this paper cites.
Diverse M-Best Solutions in Markov Random Fields
D. Batra, P. Yadollahpour, A. Guzman-Rivera, and G. Shakhnarovich · 2012
Earlier work this paper cites.
Context-dependent pre-trained deep neural networks for large-vocabulary speech recognition
G. E. Dahl, D. Yu, L. Deng, and A. Acero · 2012
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, et al · 2012
Earlier work this paper cites.
Recurrent continuous translation models
N. Kalchbrenner and P. Blunsom · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
K. Cho, B. van Merrienboer, D. Bahdanau, and Y. Bengio · 2014
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Cited alongside, same era.
A Visual Turing Test for Computer Vision Systems
D. Geman, S. Geman, N. Hallonquist, and L. Younes · 2014
Cited alongside, same era.
A multi-view embedding space for modeling internet images, tags, and their semantics
Y. Gong, Q. Ke, M. Isard, and S. Lazebnik · 2014
Cited alongside, same era.
Microsoft COCO: Common objects in context, 2014
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollar, and C. L. Zitnick · 2014
Cited alongside, same era.
A Multi-World Approach to Question Answering about Real-World Scenes based on Uncertain Input
M. Malinowski and M. Fritz · 2014
Cited alongside, same era.
Deep visual-semantic alignments for generating image descriptions
A. Karpathy and L. Fei-Fei · 2015
Later among the works it cites.
Exploring models and data for image question answering
M. Ren, R. Kiros, and R. Zemel · 2015
Later among the works it cites.
Fast and accurate recurrent neural network acoustic models for speech recognition
H. Sak, A. Senior, K. Rao, and F. Beaufays · 2015
Later among the works it cites.
Cider: Consensus-based image description evaluation
R. Vedantam, C. Lawrence Zitnick, and D. Parikh · 2015
Later among the works it cites.
O. Vinyals and Q. V. Le · 2015
Later among the works it cites.
Show and tell: A neural image caption generator
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. Simonyan and A. Zisserman · 2014
Cited alongside, same era.
Vqa: Visual question answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Cited alongside, same era.
Bidirectional recurrent neural networks as generative models
M. Berglund, T. Raiko, M. Honkala, L. Karkkainen, A. Vetek, and J. Karhunen · 2015
Cited alongside, same era.
Mind’s Eye: A Recurrent Visual Representation for Image Caption Generation
X. Chen and C. L. Zitnick · 2015
Cited alongside, same era.
Long-term Recurrent Convolutional Networks for Visual Recognition and Description
J. Donahue, L. A. Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell · 2015
Cited alongside, same era.
From Captions to Visual Concepts and Back
H. Fang, S. Gupta, F. N. Iandola, R. K. Srivastava, L. Deng, P. Dollár, J. Gao, X. He, M. Mitchell, J. C. Platt, C. L. Zitnick, and G. Zweig · 2015
Cited alongside, same era.
Bidirectional lstm-crf models for sequence tagging
Z. Huang, W. Xu, and K. Yu · 2015
Cited alongside, same era.
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2015
Later among the works it cites.
Visual Madlibs: Fill in the blank Description Generation and Question Answering
L. Yu, E. Park, A. C. Berg, and T. L. Berg · 2015
Later among the works it cites.
Visual dialog
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. Moura, D. Parikh, and D. Batra · 2016
Later among the works it cites.
Diverse beam search: Decoding diverse solutions from neural sequence models
A. K. Vijayakumar, M. Cogswell, R. R. Selvaraju, Q. Sun, S. Lee, D. J. Crandall, and D. Batra · 2016
Later among the works it cites.
Image captioning with deep bidirectional lstms
C. Wang, H. Yang, C. Bartz, and C. Meinel · 2016
Later among the works it cites.
Guesswhat?! visual object discovery through multi-modal dialogue
H. de Vries, F. Strub, S. Chandar, O. Pietquin, H. Larochelle, and A. C. Courville · 2017
Closest in time.
Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning
J. Lu, C. Xiong, D. Parikh, and R. Socher · 2017
Closest in time.