Fetching the paper…
Reading the bibliography…
We present a framework for learning to describe fine-grained visual differences between instances using attribute phrases.
Logic and conversation
H. P. Grice · 1975
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Visualizing data using t-sne
L. van der Maaten and G. Hinton · 2008
Earlier work this paper cites.
Describing objects by their attributes
A. Farhadi, I. Endres, D. Hoiem, and D. Forsyth · 2009
Earlier work this paper cites.
Automatic attribute discovery and characterization from noisy web data
T. L. Berg, A. C. Berg, and J. Shih · 2010
Earlier work this paper cites.
Attribute-centric recognition for cross-category generalization
A. Farhadi, I. Endres, and D. Hoiem · 2010
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
V. Nair and G. E. Hinton · 2010
Earlier work this paper cites.
Interactively building a discriminative vocabulary of nameable attributes
D. Parikh and K. Grauman · 2011
Earlier work this paper cites.
Relative attributes
D. Parikh and K. Grauman · 2011
Earlier work this paper cites.
Recognition using visual phrases
M. A. Sadeghi and A. Farhadi · 2011
Earlier work this paper cites.
Describing clothing by semantic attributes
H. Chen, A. Gallagher, and B. Girod · 2012
Earlier work this paper cites.
WhittleSearch: Image search with relative attribute feedback
A. Kovashka, D. Parikh, and K. Grauman · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Discovering a lexicon of parts and attributes
S. Maji · 2012
Earlier work this paper cites.
Multi-attribute spaces: Calibration for attribute fusion and similarity search
W. J. Scheirer, N. Kumar, P. N. Belhumeur, and T. E. Boult · 2012
Earlier work this paper cites.
Fine-grained visual classification of aircraft
S. Maji, E. Rahtu, J. Kannala, M. Blaschko, and A. Vedaldi · 2013
Earlier work this paper cites.
Generating expressions that refer to visible objects
M. Mitchell, K. Van Deemter, and E. Reiter · 2013
Earlier work this paper cites.
Attribute dominance: What pops out?
N. Turakhia and D. Parikh · 2013
Cited alongside, same era.
Zero-shot recognition with unreliable attributes
D. Jayaraman and K. Grauman · 2014
Cited alongside, same era.
ReferItGame: Referring to objects in photographs of natural scenes
S. Kazemzadeh, V. Ordonez, M. Matten, and T. L. Berg · 2014
Cited alongside, same era.
Attribute-based classification for zero-shot visual object categorization
C. H. Lampert, H. Nickisch, and S. Harmeling · 2014
Cited alongside, same era.
Microsoft COCO: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2015
Later among the works it cites.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. S. Corrado, A. Davis, J. Dean, M. Devin, et al · 2016
Later among the works it cites.
Reasoning About Pragmatics with Neural Listeners and Speakers
J. Andreas and D. Klein · 2016
Later among the works it cites.
Natural language object retrieval
R. Hu, H. Xu, M. Rohrbach, J. Feng, K. Saenko, and T. Darrell · 2016
Later among the works it cites.
Generation and comprehension of unambiguous object descriptions
J. Mao, J. Huang, A. Toshev, O. Camburu, A. Yuille, and K. Murphy · 2016
Later among the works it cites.
Modeling context between objects for referring expression understanding
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Understanding objects in detail with fine-grained attributes
A. Vedaldi, S. Mahendran, S. Tsogkas, S. Maji, B. Girshick, J. Kannala, E. Rahtu, I. Kokkinos, M. B. Blaschko, D. Weiss, B. Taskar, K. Simonyan, N. Saphra, and S. Mohamed · 2014
Cited alongside, same era.
Similarity comparisons for interactive fine-grained categorization
C. Wah, G. Van Horn, S. Branson, S. Maji, P. Perona, and S. Belongie · 2014
Cited alongside, same era.
Evaluation of output embeddings for fine-grained image classification
Z. Akata, S. Reed, D. Walter, H. Lee, and B. Schiele · 2015
Cited alongside, same era.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Cited alongside, same era.
Exploring nearest neighbor approaches for image captioning
J. Devlin, S. Gupta, R. Girshick, M. Mitchell, and C. L. Zitnick · 2015
Cited alongside, same era.
Long-term recurrent convolutional networks for visual recognition and description
J. Donahue, L. Anne Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell · 2015
Cited alongside, same era.
V. K. Nagaraja, V. I. Morariu, and L. S. Davis · 2016
Later among the works it cites.
Learning deep representations of fine-grained visual descriptions
S. Reed, Z. Akata, H. Lee, and B. Schiele · 2016
Later among the works it cites.
Modeling context in referring expressions
L. Yu, P. Poirson, S. Yang, A. C. Berg, and T. L. Berg · 2016
Later among the works it cites.
Yin and Yang: Balancing and answering binary visual questions
P. Zhang, Y. Goyal, D. Summers-Stay, D. Batra, and D. Parikh · 2016
Later among the works it cites.
Visual dialog
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. M. F. Moura, D. Parikh, and D. Batra · 2017
Closest in time.
Guesswhat?! visual object discovery through multi-modal dialogu
H. de Vries, F. Strub, S. Chandar, O. Pietquin, H. Larochelle, and A. Courville · 2017
Closest in time.
CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning
J. Johnson, B. Hariharan, L. van der Maaten, L. Fei-Fei, C. L. Zitnick, and R. Girshick · 2017
Closest in time.
Comprehension-guided referring expressions
R. Luo and G. Shakhnarovich · 2017
Closest in time.
A taxonomy of part and attribute discovery techniques
S. Maji · 2017
Closest in time.
Context-aware captions from context-agnostic supervision
R. Vedantam, S. Bengio, K. Murphy, D. Parikh, and G. Chechik · 2017
Closest in time.
A joint speaker-listener-reinforcer model for referring expressions
L. Yu, H. Tan, M. Bansal, and T. L. Berg · 2017
Closest in time.