Fetching the paper…
Reading the bibliography…
In this paper, we study the task of image retrieval, where the input query is specified in the form of an image plus some text that describes desired modifications to the input image.
Relevance feedback: a power tool for interactive content-based image retrieval
Y. Rui, T. S. Huang, M. Ortega, and S. Mehrotra · 1998
Earlier work this paper cites.
Learning a similarity metric discriminatively, with application to face verification
S. Chopra, R. Hadsell, and Y. LeCun · 2005
Earlier work this paper cites.
Neighbourhood components analysis
J. Goldberger, G. E. Hinton, S. T. Roweis, and R. R. Salakhutdinov · 2005
Earlier work this paper cites.
Im2gps: estimating geographic information from a single image
J. Hays and A. A. Efros · 2008
Earlier work this paper cites.
Describing objects by their attributes
A. Farhadi, I. Endres, D. Hoiem, and D. Forsyth · 2009
Earlier work this paper cites.
Learning to detect unseen object classes by between-class attribute transfer
C. H. Lampert, H. Nickisch, and S. Harmeling · 2009
Earlier work this paper cites.
Recognition using visual phrases
M. A. Sadeghi and A. Farhadi · 2011
Earlier work this paper cites.
Leveraging high-level and low-level features for multimedia event detection
L. Jiang, A. G. Hauptmann, and G. Xiang · 2012
Earlier work this paper cites.
Whittlesearch: Image search with relative attribute feedback
A. Kovashka, D. Parikh, and K. Grauman · 2012
Earlier work this paper cites.
Inferring analogous attributes
C.-Y. Chen and K. Grauman · 2014
Earlier work this paper cites.
Learning fine-grained image similarity with deep ranking
J. Wang, Y. Song, T. Leung, C. Rosenberg, J. Wang, J. Philbin, B. Chen, and Y. Wu · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Discovering states and transformations in image collections
P. Isola, J. J. Lim, and E. H. Adelson · 2015
Earlier work this paper cites.
Bridging the ultimate semantic gap: A semantic search engine for internet videos
L. Jiang, S.-I. Yu, D. Meng, T. Mitamura, and A. G. Hauptmann · 2015
Earlier work this paper cites.
Learning deep representations for ground-to-aerial geolocalization
T.-Y. Lin, Y. Cui, S. Belongie, and J. Hays · 2015
Cited alongside, same era.
Deep face recognition
O. M. Parkhi, A. Vedaldi, A. Zisserman, et al · 2015
Cited alongside, same era.
An embarrassingly simple approach to zero-shot learning
B. Romera-Paredes and P. Torr · 2015
Cited alongside, same era.
Facenet: A unified embedding for face recognition and clustering
F. Schroff, D. Kalenichenko, and J. Philbin · 2015
Cited alongside, same era.
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2015
Cited alongside, same era.
Zero-shot learning via semantic similarity embedding
Z. Zhang and V. Saligrama · 2015
Cited alongside, same era.
In defense of the triplet loss for person re-identification
A. Hermans, L. Beyer, and B. Leibe · 2017
Later among the works it cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
J. Johnson, B. Hariharan, L. van der Maaten, L. Fei-Fei, C. L. Zitnick, and R. Girshick · 2017
Later among the works it cites.
From red wine to red tomato: Composition with context
I. Misra, A. Gupta, and M. Hebert · 2017
Later among the works it cites.
No fuss distance metric learning using proxies
Y. Movshovitz-Attias, A. Toshev, T. K. Leung, S. Ioffe, and S. Singh · 2017
Later among the works it cites.
A simple neural network module for relational reasoning
A. Santoro, D. Raposo, D. G. Barrett, M. Malinowski, R. Pascanu, P. Battaglia, and T. Lillicrap · 2017
Later among the works it cites.
Prototypical networks for few-shot learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep image retrieval: Learning global representations for image search
A. Gordo, J. Almazán, J. Revaud, and D. Larlus · 2016
Cited alongside, same era.
Deepfashion: Powering robust clothes recognition and retrieval with rich annotations
Z. Liu, P. Luo, S. Qiu, X. Wang, and X. Tang · 2016
Cited alongside, same era.
Image question answering using convolutional neural network with dynamic parameter prediction
H. Noh, P. Hongsuck Seo, and B. Han · 2016
Cited alongside, same era.
Cnn image retrieval learns from bow: Unsupervised fine-tuning with hard examples
F. Radenović, G. Tolias, and O. Chum · 2016
Cited alongside, same era.
The sketchy database: learning to retrieve badly drawn bunnies
P. Sangkloy, N. Burnell, C. Ham, and J. Hays · 2016
Cited alongside, same era.
Localizing and orienting street views using overhead imagery
N. N. Vo and J. Hays · 2016
Cited alongside, same era.
J. Snell, K. Swersky, and R. Zemel · 2017
Later among the works it cites.
Memory-augmented attribute manipulation networks for interactive fashion search
B. Zhao, J. Feng, X. Wu, and S. Yan · 2017
Later among the works it cites.
Learning attribute representations with localization for flexible fashion search
K. E. Ak, A. A. Kassim, J. H. Lim, and J. Y. Tham · 2018
Closest in time.
Dynamic few-shot visual learning without forgetting
S. Gidaris and N. Komodakis · 2018
Closest in time.
Dialog-based interactive image retrieval
X. Guo, H. Wu, Y. Cheng, S. Rennie, and R. S. Feris · 2018
Closest in time.
Compositional learning for human object interaction
K. Kato, Y. Li, and A. Gupta · 2018
Closest in time.
Focal visual-text attention for visual question answering
J. Liang, L. Jiang, L. Cao, L.-J. Li, and A. Hauptmann · 2018
Closest in time.
Attributes as operators
T. Nagarajan and K. Grauman · 2018
Closest in time.
Film: Visual reasoning with a general conditioning layer
E. Perez, F. Strub, H. De Vries, V. Dumoulin, and A. Courville · 2018
Closest in time.