Fetching the paper…
Reading the bibliography…
We address the problem of recognizing situations in images.
Generalization of backpropagation with application to a recurrent gas market model
P. J. Werbos · 1988
Earlier work this paper cites.
WordNet: A Lexical Database for English
G. A. Miller · 1995
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
From TreeBank to PropBank
P. Kingsbury and M. Palmer · 2002
Earlier work this paper cites.
Background to FrameNet
C. J. Fillmore, C. R. Johnson, and M. R. L. Petruck · 2003
Earlier work this paper cites.
Beyond Nouns: Exploiting Prepositions and Comparative Adjectives for Learning Visual Classifiers
A. Gupta and L. S. Davis · 2008
Earlier work this paper cites.
Visualizing High-Dimensional Data using t-SNE
L. J. P. van der Maaten and G. E. Hinton · 2008
Earlier work this paper cites.
Graph Alignment for Semi-Supervised Semantic Role Labeling
H. Fürstenau and M. Lapata · 2009
Earlier work this paper cites.
The Graph Neural Network Model
F. Scarselli, M. Gori, A. C. Tsoi, M. Hagenbuchner, and G. Monfardini · 2009
Earlier work this paper cites.
Recognizing human actions in still images: a study of bag-of-features and part-based representations
V. Delaitre, I. Laptev, and J. Sivic · 2010
Earlier work this paper cites.
Grouplet: A Structured Image Representation for Recognizing Human and Object Interactions
B. Yao and L. Fei-Fei · 2010
Earlier work this paper cites.
Torch7: A matlab-like environment for machine learning
R. Collobert, K. Kavukcuoglu, and C. Farabet · 2011
Earlier work this paper cites.
Action recognition by dense trajectories
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu · 2011
Earlier work this paper cites.
Human Action Recognition by Learning Bases of Action Attributes and Parts
B. Yao, X. Jiang, A. Khosla, A. L. Lin, L. Guibas, and L. Fei-Fei · 2011
Earlier work this paper cites.
Spectral networks and locally connected networks on graphs
J. Bruna, W. Zaremba, A. Szlam, and Y. LeCun · 2014
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
K. Cho, B. van Merriënboer, C. Gulcehre, B. Dzmitry, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Learning Deep Features for Scene Recognition using Places Database
B. Zhou, A. Lapedriza, J. Xiao, A. Torralba, and A. Oliva · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Cited alongside, same era.
Convolutional networks on graphs for learning molecular fingerprints
D. K. Duvenaud, D. Maclaurin, J. Iparraguirre, R. Bombarell, T. Hirzel, A. Aspuru-Guzik, and R. P. Adams · 2015
Cited alongside, same era.
S. Gupta and J. Malik · 2015
Cited alongside, same era.
Image Retrieval using Scene Graphs
J. Johnson, R. Krishna, M. Stark, L.-J. Li, D. A. Shamma, M. S. Bernstein, and L. Fei-Fei · 2015
Cited alongside, same era.
Deep Visual-Semantic Alignments for Generating Image Descriptions
A. Karpathy and L. Fei-Fei · 2015
Cited alongside, same era.
Generating multi-sentence lingual descriptions of indoor scenes
D. Lin, S. Fidler, C. Kong, and R. Urtasun · 2015
Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma, M. Bernstein, and L. Fei-Fei · 2016
Later among the works it cites.
Gated Graph Sequence Neural Networks
Y. Li, D. Tarlow, M. Brockschmidt, and R. Zemel · 2016
Later among the works it cites.
Semantic object parsing with graph lstm
X. Liang, X. Shen, J. Feng, L. Lin, and S. Yan · 2016
Later among the works it cites.
Visual Relationship Detection with Language Priors
C. Lu, R. Krishna, M. Bernstein, and L. Fei-Fei · 2016
Later among the works it cites.
Neural Semantic Role Labeling with Dependency Path Embeddings
M. Roth and M. Lapata · 2016
Later among the works it cites.
MovieQA: Understanding Stories in Movies through Question-Answering
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Describing Common Human Visual Actions in Images
M. R. Ronchi and P. Perona · 2015
Cited alongside, same era.
ImageNet Large Scale Visual Recognition Challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Cited alongside, same era.
Improved semantic representations from tree-structured long short-term memory networks
K. S. Tai, R. Socher, and C. D. Manning · 2015
Cited alongside, same era.
Show and Tell: A Neural Image Caption Generator
O. Vinyals, A. Toshev, S. Bengio, and Dumitru Erhan · 2015
Cited alongside, same era.
Show, Attend and Tell: Neural Image Caption Generation with Visual Attention
K. Xu, J. L. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhutdinov, R. S. Zemel, and Y. Bengio · 2015
Cited alongside, same era.
M. Tapaswi, Y. Zhu, R. Stiefelhagen, A. Torralba, R. Urtasun, and S. Fidler · 2016
Later among the works it cites.
Grounded Semantic Role Labeling
S. Yang, Q. Gao, C. Liu, C. Xiong, S.-C. Zhu, and J. Y. Chai · 2016
Later among the works it cites.
Situation Recognition: Visual Semantic Role Labeling for Image Understanding
M. Yatskar, L. Zettlemoyer, and A. Farhadi · 2016
Later among the works it cites.
Places: An Image Database for Deep Scene Understanding
B. Zhou, A. Khosla, A. Lapedriza, A. Torralba, and A. Oliva · 2016
Later among the works it cites.
Annotating object instances with a polygon-rnn
L. Castrejon, K. Kundu, R. Urtasun, and S. Fidler · 2017
Closest in time.
Speech and Language Processing
D. Jurafsky and J. H. Martin · 2017
Closest in time.
Semi-supervised classification with graph convolutional networks
T. N. Kipf and M. Welling · 2017
Closest in time.
Teaching machines to describe images via natural language feedback
H. Ling and S. Fidler · 2017
Closest in time.
The more you know: Using knowledge graphs for image classification
K. Marino, R. Salakhutdinov, and A. Gupta · 2017
Closest in time.
Phrase Localization and Visual Relationship Detection with Comprehension Linguistic Cues
B. A. Plummer, A. Mallya, C. M. Cervantes, and J. Hockenmaier · 2017
Closest in time.
Commonly Uncommon: Semantic Sparsity in Situation Recognition
M. Yatskar, V. Ordonez, L. Zettlemoyer, and A. Farhadi · 2017
Closest in time.
Visual Translation Embedding Network for Visual Relation Detection
H. Zhang, Z. Kyaw, S.-F. Chang, and T.-S. Chua · 2017
Closest in time.