Fetching the paper…
Reading the bibliography…
In this paper, we address referring expression comprehension: localizing an image region described by a natural language expression.
Parsing with compositional vector grammars
R. Socher, J. Bauer, C. D. Manning, et al · 2013
Earlier work this paper cites.
Referitgame: Referring to objects in photographs of natural scenes
S. Kazemzadeh, V. Ordonez, M. Matten, and T. L. Berg · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Earlier work this paper cites.
Learning to compose neural networks for question answering
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Earlier work this paper cites.
Neural module networks
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Segmentation from natural language expressions
R. Hu, M. Rohrbach, and T. Darrell · 2016
Earlier work this paper cites.
Natural language object retrieval
R. Hu, H. Xu, M. Rohrbach, J. Feng, K. Saenko, and T. Darrell · 2016
Earlier work this paper cites.
Generation and comprehension of unambiguous object descriptions
J. Mao, J. Huang, A. Toshev, O. Camburu, A. Yuille, and K. Murphy · 2016
Cited alongside, same era.
Modeling context between objects for referring expression understanding
V. K. Nagaraja, V. I. Morariu, and L. S. Davis · 2016
Cited alongside, same era.
Grounding of textual phrases in images by reconstruction
A. Rohrbach, M. Rohrbach, R. Hu, T. Darrell, and B. Schiele · 2016
Cited alongside, same era.
Learning deep structure-preserving image-text embeddings
L. Wang, Y. Li, and S. Lazebnik · 2016
Cited alongside, same era.
Hierarchical attention networks for document classification
Z. Yang, D. Yang, C. Dyer, X. He, A. J. Smola, and E. H. Hovy · 2016
Cited alongside, same era.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Later among the works it cites.
Learning to reason: End-to-end module networks for visual question answering
R. Hu, J. Andreas, M. Rohrbach, T. Darrell, and K. Saenko · 2017
Later among the works it cites.
Modeling relationship in referential expressions with compositional modular networks
R. Hu, M. Rohrbacnh, J. Andreas, T. Darrell, and K. Saenko · 2017
Later among the works it cites.
Inferring and executing programs for visual reasoning
J. Johnson, B. Hariharan, L. van der Maaten, J. Hoffman, L. Fei-Fei, C. L. Zitnick, and R. Girshick · 2017
Later among the works it cites.
Recurrent multimodal interaction for referring image segmentation
C. Liu, Z. Lin, X. Shen, J. Yang, X. Lu, and A. Yuille · 2017
Later among the works it cites.
Referring expression generation and comprehension via attributes
J. Liu, L. Wang, and M.-H. Yang · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Yao, Y. Pan, Y. Li, Z. Qiu, and T. Mei · 2016
Cited alongside, same era.
Image captioning with semantic attention
Q. You, H. Jin, Z. Wang, C. Fang, and J. Luo · 2016
Cited alongside, same era.
Modeling context in referring expressions
L. Yu, P. Poirson, S. Yang, A. C. Berg, and T. L. Berg · 2016
Cited alongside, same era.
Modular multitask reinforcement learning with policy sketches
J. Andreas, D. Klein, and S. Levine · 2017
Cited alongside, same era.
Query-guided regression network with context policy for phrase grounding
K. Chen, R. Kovvuri, and R. Nevatia · 2017
Cited alongside, same era.
An implementation of faster rcnn with study for region sampling
X. Chen and A. Gupta · 2017
Cited alongside, same era.
Later among the works it cites.
Comprehension-guided referring expressions
R. Luo and G. Shakhnarovich · 2017
Later among the works it cites.
Reasoning about fine-grained attribute phrases using reference games
J.-C. Su, C. Wu, H. Jiang, and S. Maji · 2017
Later among the works it cites.
Learning two-branch neural networks for image-text matching tasks
L. Wang, Y. Li, and S. Lazebnik · 2017
Later among the works it cites.
Image captioning and visual question answering based on attributes and external knowledge
Q. Wu, C. Shen, P. Wang, A. Dick, and A. van den Hengel · 2017
Later among the works it cites.
A joint speaker-listener-reinforcer model for referring expressions
L. Yu, H. Tan, M. Bansal, and T. L. Berg · 2017
Later among the works it cites.