V. I. Levenshtein, “Binary Codes Capable of Correcting Deletions, Insertions and Reversals,” Soviet Physics Doklady , vol. 10, p. 707, Feb. 1966
1966
Earlier work this paper cites.
M. Grubinger, P. Clough, H. Müller, and T. Deselaers, “The IAPR benchmark: A new evaluation resource for visual information systems,” in ICLRE , 2006
2006
Earlier work this paper cites.
R. Hadsell, S. Chopra, and Y. LeCun, “Dimensionality reduction by learning an invariant mapping,” in CVPR , 2006
2006
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “ImageNet: a large-scale hierarchical image database,” in CVPR , 2009
2009
Earlier work this paper cites.
T.-S. Chua, J. Tang, R. Hong, H. Li, Z. Luo, and Y. Zheng, “NUS-WIDE: A real-world web image database from National University of Singapore,” in CIVR , 2009
2009
Earlier work this paper cites.
C. Rashtchian, P. Young, M. Hodosh, and J. Hockenmaier, “Collecting image annotations using Amazon’s Mechanical Turk,” in NAACL-HLT Workshop , 2010
2010
Earlier work this paper cites.
Y. Peng, J. Qi, and Y. Yuan, “Modality-specific cross-modal similarity measurement with recurrent attention network,” IEEE Trans. Image Processing , vol. 27, no. 11, pp. 5585–5599, 2018
2012
Earlier work this paper cites.
B. Caputo, H. Müller, B. Thomee, M. Villegas, R. Paredes, D. Zellhofer, H. Goeau, A. Joly, P. Bonnet, J. Gomez, I. Varea, and M. Cazorla, “ImageCLEF 2013: The vision, the data and the open challenges,” in CLEF , 2013
2013
Earlier work this paper cites.
M. Hodosh, P. Young, and J. Hockenmaier, “Framing image description as a ranking task: Data, models and evaluation metrics,” JAIR , vol. 47, pp. 853–899, 2013
2013
Earlier work this paper cites.
P. Young, A. Lai, M. Hodosh, and J. Hockenmaier, “From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,” TACL , vol. 2, pp. 67–78, 2014
2014
Earlier work this paper cites.
Z. Wang, P. Cui, L. Xie, W. Zhu, Y. Rui, and S. Yang, “Bilateral correspondence model for words-and-pictures association in multimedia-rich microblogs,” ACM Trans. Multimedia Comput. Commun. Appl. , vol. 10, no. 4, pp. 34:1–34:21, 2014
2014
Earlier work this paper cites.
P. Cui, S.-W. Liu, W.-W. Zhu, H.-B. Luan, T.-S. Chua, and S.-Q. Yang, “Social-sensed image search,” ACM Trans. Inf. Syst. , vol. 32, no. 2, pp. 8:1–8:23, 2014
2014
Earlier work this paper cites.
X. Chen, H. Fang, T.-Y. Lin, R. Vedantam, S. Gupta, P. Dollár, and L. Zitnick, “Microsoft COCO captions: Data collection and evaluation server,” CoRR , vol. abs/1504.00325, 2015
Original
2015
Earlier work this paper cites.
R. Funaki and H. Nakayama, “Image-mediated learning for zero-shot cross-lingual document retrieval.” in EMNLP , 2015
2015
Earlier work this paper cites.
K. Min, C. Ma, T. Zhao, and H. Li, “BosonNLP: An ensemble approach for word segmentation and POS tagging,” in NLPCC , 2015
2015
Earlier work this paper cites.
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan, “Show and tell: A neural image caption generator,” in CVPR , 2015
2015
Earlier work this paper cites.
D. Bahdanau, K. Cho, and Y. Bengio, “Neural machine translation by jointly learning to align and translate,” in ICLR , 2015
2015
Earlier work this paper cites.