Fetching the paper…
Reading the bibliography…
Significant progress has been achieved in Computer Vision by leveraging large-scale image datasets.
Edgeflow: a technique for boundary detection and image segmentation
W.-Y. Ma and B. S. Manjunath · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
C.-Y. Lin · 2004
Earlier work this paper cites.
Describing objects by their attributes
A. Farhadi, I. Endres, D. Hoiem, and D. Forsyth · 2009
Earlier work this paper cites.
Zero-shot learning with semantic output codes
M. Palatucci, D. Pomerleau, G. E. Hinton, and T. M. Mitchell · 2009
Earlier work this paper cites.
Attribute-based people search in surveillance environments
D. A. Vaquero, R. S. Feris, D. Tran, L. Brown, A. Hampapur, and M. Turk · 2009
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
M. Everingham, L. Gool, C. K. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Every picture tells a story: Generating sentences from images
A. Farhadi, M. Hejrati, M. A. Sadeghi, P. Young, C. Rashtchian, J. Hockenmaier, and D. Forsyth · 2010
Earlier work this paper cites.
Attribute learning in large-scale datasets
O. Russakovsky and F.-F. Li · 2010
Earlier work this paper cites.
Sun database: Large-scale scene recognition from abbey to zoo
J. Xiao, J. Hays, K. A. Ehinger, A. Oliva, and A. Torralba · 2010
Earlier work this paper cites.
Attribute-based transfer learning for object categorization with zero/one training example
X. Yu and Y. Aloimonos · 2010
Earlier work this paper cites.
Im2text: Describing images using 1 million captioned photographs
V. Ordonez, G. Kulkarni, and T. L. Berg · 2011
Earlier work this paper cites.
Articulated part-based model for joint object detection and pose estimation
M. Sun and S. Savarese · 2011
Earlier work this paper cites.
The Caltech-UCSD Birds-200-2011 Dataset
C. Wah, S. Branson, P. Welinder, P. Perona, and S. Belongie · 2011
Earlier work this paper cites.
Corpus-guided sentence generation of natural images
Y. Yang, C. L. Teo, H. Daumé III, and Y. Aloimonos · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Collective generation of natural image descriptions
P. Kuznetsova, V. Ordonez, A. C. Berg, T. L. Berg, and Y. Choi · 2012
Earlier work this paper cites.
Articulated people detection and pose estimation: Reshaping the future
L. Pishchulin, A. Jain, M. Andriluka, and T. Thormahlen · 2012
Earlier work this paper cites.
Babytalk: Understanding and generating simple image descriptions
G. Kulkarni, V. Premraj, V. Ordonez, S. Dhar, S. Li, Y. Choi, A. C. Berg, and T. L. Berg · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean · 2013
Earlier work this paper cites.
2d human pose estimation: New benchmark and state of the art analysis
M. Andriluka, L. Pishchulin, P. Gehler, and B. Schiele · 2014
Earlier work this paper cites.
Articulated pose estimation by a graphical model with image dependent pairwise relations
X. Chen and A. L. Yuille · 2014
Earlier work this paper cites.
Meteor universal: Language specific translation evaluation for any target language
M. Denkowski and A. Lavie · 2014
Cited alongside, same era.
Using k-poselets for detecting people and localizing their keypoints
G. Gkioxari, B. Hariharan, R. Girshick, and J. Malik · 2014
Cited alongside, same era.
Attribute-based classification for zero-shot visual object categorization
C. H. Lampert, H. Nickisch, and S. Harmeling · 2014
Cited alongside, same era.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Cited alongside, same era.
Multi-source deep learning for human pose estimation
W. Ouyang, X. Chu, and X. Wang · 2014
Cited alongside, same era.
Joint training of a convolutional network and a graphical model for human pose estimation
J. Tompson, A. Jain, Y. Lecun, C. Bregler, J. Tompson, A. Jain, Y. Lecun, and C. Bregler · 2014
Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs
L. C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille · 2016
Later among the works it cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Later among the works it cites.
Multi-person pose estimation with local joint-to-person associations
U. Iqbal and J. Gall · 2016
Later among the works it cites.
Adding chinese captions to images
X. Li, W. Lan, J. Dong, and H. Liu · 2016
Later among the works it cites.
Stacked Hourglass Networks for Human Pose Estimation
A. Newell, K. Yang, and J. Deng · 2016
Later among the works it cites.
Convolutional pose machines
S. E. Wei, V. Ramakrishna, T. Kanade, and Y. Sheikh · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deeppose: Human pose estimation via deep neural networks
A. Toshev and C. Szegedy · 2014
Cited alongside, same era.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
P. Young, A. Lai, M. Hodosh, and J. Hockenmaier · 2014
Cited alongside, same era.
Microsoft coco captions: Data collection and evaluation server
X. Chen, H. Fang, T.-Y. Lin, R. Vedantam, S. Gupta, P. Dollár, and C. L. Zitnick · 2015
Cited alongside, same era.
Language models for image captioning: The quirks and what works
J. Devlin, H. Cheng, H. Fang, S. Gupta, L. Deng, X. He, G. Zweig, and M. Mitchell · 2015
Cited alongside, same era.
From captions to visual concepts and back
H. Fang, S. Gupta, F. Iandola, R. K. Srivastava, L. Deng, P. Dollar, J. Gao, X. He, M. Mitchell, J. C. Platt, C. Lawrence Zitnick, and G. Zweig · 2015
Cited alongside, same era.
Framing image description as a ranking task: Data, models and evaluation metrics
M. Hodosh, P. Young, and J. Hockenmaier · 2015
Cited alongside, same era.
Latent embeddings for zero-shot classification
Y. Xian, Z. Akata, G. Sharma, Q. Nguyen, M. Hein, and B. Schiele · 2016
Later among the works it cites.
Boosting image captioning with attributes
T. Yao, Y. Pan, Y. Li, Z. Qiu, and T. Mei · 2016
Later among the works it cites.
Image captioning with semantic attention
Q. You, H. Jin, Z. Wang, C. Fang, and J. Luo · 2016
Later among the works it cites.
Zero-shot recognition via structured prediction
Z. Zhang and V. Saligrama · 2016
Later among the works it cites.
Sca-cnn: Spatial and channel-wise attention in convolutional networks for image captioning
L. Chen, H. Zhang, J. Xiao, L. Nie, J. Shao, W. Liu, and T.-S. Chua · 2017
Closest in time.
Recent advances in zero-shot recognition
Y. Fu, T. Xiang, Y.-G. Jiang, X. Xue, L. Sigal, and S. Gong · 2017
Closest in time.
K. He, G. Gkioxari, P. Dollár, and R. B. Girshick · 2017
Closest in time.
Gaze embeddings for zero-shot image classification
N. Karessli, Z. Akata, B. Schiele, and A. Bulling · 2017
Closest in time.
Semantic autoencoder for zero-shot learning
E. Kodirov, T. Xiang, and S. Gong · 2017
Closest in time.
Deep reinforcement learning-based image captioning with embedding reward
Z. Ren, X. Wang, N. Zhang, X. Lv, and L.-J. Li · 2017
Closest in time.
Image captioning and visual question answering based on attributes and external knowledge
Q. Wu, C. Shen, P. Wang, A. Dick, and A. van den Hengel · 2017
Closest in time.
Zero-shot learning-a comprehensive evaluation of the good, the bad and the ugly
Y. Xian, C. H. Lampert, B. Schiele, and Z. Akata · 2017
Closest in time.
Zero-shot classification with discriminative semantic representation learning
M. Ye and Y. Guo · 2017
Closest in time.
Learning a deep embedding model for zero-shot learning
L. Zhang, T. Xiang, and S. Gong · 2017
Closest in time.
Zero-shot learning posed as a missing data problem
B. Zhao, B. Wu, T. Wu, and Y. Wang · 2017
Closest in time.