Fetching the paper…
Reading the bibliography…
This paper proposes a new evaluation protocol for cross-media retrieval which better fits the real-word applications.
H. Hotelling, “Relations between two sets of variates,” Biometrika , vol. 28, no. 3/4, pp. 321–377, 1936
1936
Earlier work this paper cites.
L. Zheng, S. Wang, Z. Liu, and Q. Tian, “Packing and padding: Coupled multi-index for accurate image retrieval,” in CVPR , 2014, pp. 1939–1946
1946
Earlier work this paper cites.
M. Turk and A. Pentland, “Eigenfaces for recognition,” JOCN , vol. 3, no. 1, pp. 71–86, 1991
1991
Earlier work this paper cites.
P. N. Belhumeur, J. P. Hespanha, and D. J. Kriegman, “Eigenfaces vs. fisherfaces: Recognition using class specific linear projection,” TPAMI , vol. 19, no. 7, pp. 711–720, 1997
1997
Earlier work this paper cites.
F. Wu, Y. Yang, Y. Zhuang, and Y. Pan, “Understanding multimedia document semantics for cross-media retrieval,” in PCM . Springer Berlin Heidelberg, 2005, pp. 993–1004
2005
Earlier work this paper cites.
S. Yan, D. Xu, B. Zhang, H.-J. Zhang, Q. Yang, and S. Lin, “Graph embedding and extensions: A general framework for dimensionality reduction,” TPAMI , vol. 29, no. 1, 2007
2007
Earlier work this paper cites.
Y.-T. Zhuang, Y. Yang, and F. Wu, “Mining semantic correlation of heterogeneous multimedia data for cross-media retrieval,” TMM , vol. 10, no. 2, pp. 221–229, 2008
2008
Earlier work this paper cites.
Y. Yang, Y.-T. Zhuang, F. Wu, and Y.-H. Pan, “Harmonizing hierarchical manifolds for multimedia document semantics understanding and cross-media retrieval,” TMM , vol. 10, no. 3, pp. 437–446, 2008
2008
Earlier work this paper cites.
M. J. Huiskes and M. S. Lew, “The mir flickr retrieval evaluation,” in MIR . ACM, 2008, pp. 39–43
2008
Earlier work this paper cites.
Y. Yang, D. Xu, F. Nie, J. Luo, and Y. Zhuang, “Ranking with local regression and global alignment for cross media retrieval,” in ACMMM . ACM, 2009, pp. 175–184
2009
Earlier work this paper cites.
T.-S. Chua, J. Tang, R. Hong, H. Li, Z. Luo, and Y. Zheng, “Nus-wide: a real-world web image database from national university of singapore,” in CIVR . ACM, 2009, pp. 1–9
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in CVPR . IEEE, 2009, pp. 248–255
2009
Earlier work this paper cites.
Y. Weiss, A. Torralba, and R. Fergus, “Spectral hashing,” in NIPS , 2009, pp. 1753–1760
2009
Earlier work this paper cites.
Y. Yang, F. Wu, D. Xu, Y. Zhuang, and L.-T. Chia, “Cross-media retrieval using query dependent search methods,” Pattern Recognition , vol. 43, no. 8, pp. 2927–2936, 2010
2010
Earlier work this paper cites.
N. Rasiwasia, J. Costa Pereira, E. Coviello, G. Doyle, G. R. Lanckriet, R. Levy, and N. Vasconcelos, “A new approach to cross-modal multimedia retrieval,” in ACMMM . ACM, 2010, pp. 251–260
2010
Earlier work this paper cites.
M. M. Bronstein, A. M. Bronstein, F. Michel, and N. Paragios, “Data fusion through cross-modality metric learning using similarity-sensitive hashing,” in CVPR . IEEE, 2010, pp. 3594–3601
2010
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes (voc) challenge,” IJCV , vol. 88, no. 2, pp. 303–338, 2010
2010
Earlier work this paper cites.
C. Rashtchian, P. Young, M. Hodosh, and J. Hockenmaier, “Collecting image annotations using amazon’s mechanical turk,” in NAACL Workshop . Association for Computational Linguistics, 2010, pp. 139–147
2010
Earlier work this paper cites.
S. Kumar and R. Udupa, “Learning hash functions for cross-view similarity search,” in IJCAI , vol. 22, no. 1, 2011, pp. 1360–1365
2011
Earlier work this paper cites.
V. Ordonez, G. Kulkarni, and T. L. Berg, “Im2text: Describing images using 1 million captioned photographs,” in NIPS , 2011, pp. 1143–1151
2011
Earlier work this paper cites.
A. Sharma and D. W. Jacobs, “Bypassing synthesis: Pls for face recognition with pose, low-resolution and sketch,” in CVPR . IEEE, 2011, pp. 593–600
2011
Earlier work this paper cites.
A. Sharma, A. Kumar, H. Daume, and D. W. Jacobs, “Generalized multiview analysis: A discriminative latent space,” in CVPR . IEEE, 2012, pp. 2160–2167
2012
Earlier work this paper cites.
Y. Zhen and D.-Y. Yeung, “A probabilistic model for multimodal hash function learning,” in SIGKDD . ACM, 2012, pp. 940–948
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in NIPS , 2012, pp. 1097–1105
2012
Cited alongside, same era.
J. Song, Y. Yang, Z. Huang, H. T. Shen, and J. Luo, “Effective multiple feature hashing for large-scale near-duplicate video retrieval,” TMM , vol. 15, no. 8, pp. 1997–2008, 2013
2013
Cited alongside, same era.
G. Andrew, R. Arora, J. A. Bilmes, and K. Livescu, “Deep canonical correlation analysis.” in ICML , 2013, pp. 1247–1255
2013
Cited alongside, same era.
J. Song, Y. Yang, Y. Yang, Z. Huang, and H. T. Shen, “Inter-media hashing for large-scale retrieval from heterogeneous data sources,” in SIGMOD . ACM, 2013, pp. 785–796
2013
Cited alongside, same era.
A. Frome, G. S. Corrado, J. Shlens, S. Bengio, J. Dean, T. Mikolov et al. , “Devise: A deep visual-semantic embedding model,” in NIPS , 2013, pp. 2121–2129
Z. Lin, G. Ding, M. Hu, and J. Wang, “Semantics-preserving hashing for cross-view retrieval,” in CVPR , 2015, pp. 3864–3872
2015
Later among the works it cites.
L. Ma, Z. Lu, L. Shang, and H. Li, “Multimodal convolutional neural networks for matching image and sentence,” in ICCV , 2015, pp. 2623–2631
2015
Later among the works it cites.
X. Chen and C. Lawrence Zitnick, “Mind’s eye: A recurrent visual representation for image caption generation,” in CVPR , 2015, pp. 2422–2431
2015
Later among the works it cites.
A. Karpathy and L. Fei-Fei, “Deep visual-semantic alignments for generating image descriptions,” in CVPR , 2015, pp. 3128–3137
2015
Later among the works it cites.
L. Zheng, L. Shen, L. Tian, S. Wang, J. Wang, and Q. Tian, “Scalable person re-identification: A benchmark,” in ICCV , 2015, pp. 1116–1124
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
M. Hodosh, P. Young, and J. Hockenmaier, “Framing image description as a ranking task: Data, models and evaluation metrics,” JAIR , vol. 47, pp. 853–899, 2013
2013
Cited alongside, same era.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in NIPS , 2013, pp. 3111–3119
2013
Cited alongside, same era.
Y. Gong, Q. Ke, M. Isard, and S. Lazebnik, “A multi-view embedding space for modeling internet images, tags, and their semantics,” IJCV , vol. 106, no. 2, pp. 210–233, 2014
2014
Cited alongside, same era.
N. Rasiwasia, D. Mahajan, V. Mahadevan, and G. Aggarwal, “Cluster canonical correlation analysis.” in AISTATS , 2014, pp. 823–831
2014
Cited alongside, same era.
R. Socher, A. Karpathy, Q. V. Le, C. D. Manning, and A. Y. Ng, “Grounded compositional semantics for finding and describing images with sentences,” TACL , vol. 2, pp. 207–218, 2014
2014
Cited alongside, same era.
2014
Cited alongside, same era.
F. Feng, X. Wang, and R. Li, “Cross-modal retrieval with correspondence autoencoder,” in ACMMM . ACM, 2014, pp. 7–16
2014
Cited alongside, same era.
L. Zheng, S. Wang, J. Wang, and Q. Tian, “Accurate image search with multi-scale contextual evidences,” IJCV , vol. 120, no. 1, pp. 1–13, 2016
2016
Later among the works it cites.
2016
Later among the works it cites.
L. Wang, Y. Li, and S. Lazebnik, “Learning deep structure-preserving image-text embeddings,” in CVPR , 2016, pp. 5005–5013
2016
Later among the works it cites.
Y. Wei, Y. Zhao, C. Lu, S. Wei, L. Liu, Z. Zhu, and S. Yan, “Cross-modal retrieval with cnn visual features: A new baseline,” IEEE Trans. on Cybernetics , 2016, Preprint
2016
Later among the works it cites.
Y. He, S. Xiang, C. Kang, J. Wang, and C. Pan, “Cross-modal retrieval via deep and bidirectional representation learning,” TMM , vol. 18, no. 7, pp. 1363–1377, 2016
2016
Later among the works it cites.
V. Vukotić, C. Raymond, and G. Gravier, “Bidirectional joint representation learning with symmetrical deep neural networks for multimodal and crossmodal applications,” in ICMR . ACM, 2016, pp. 343–346
2016
Later among the works it cites.
2016
Later among the works it cites.
X. Xu, F. Shen, Y. Yang, and H. T. Shen, “Discriminant cross-modal hashing,” in ICMR . ACM, 2016, pp. 305–308
2016
Later among the works it cites.
Q.-Y. Jiang and W.-J. Li, “Deep cross-modal hashing,” arXiv preprint arXiv:1602.02255 , 2016
2016
Later among the works it cites.
Y. Cao, M. Long, J. Wang, Q. Yang, and P. S. Yu, “Deep visual-semantic hashing for cross-modal retrieval,” in SIGKDD . ACM, 2016, pp. 1445–1454
2016
Later among the works it cites.
Y. Yan, F. Nie, W. Li, C. Gao, Y. Yang, and D. Xu, “Image classification by cross-media active learning with privileged information,” TMM , vol. 18, no. 12, pp. 2494–2502, 2016
2016
Later among the works it cites.
2016
Later among the works it cites.
F. Radenović, G. Tolias, and O. Chum, “Cnn image retrieval learns from bow: Unsupervised fine-tuning with hard examples,” in ECCV . Springer, 2016, pp. 3–20
2016
Later among the works it cites.
2016
Later among the works it cites.
X. Liu, W. Liu, T. Mei, and H. Ma, “A deep learning-based approach to progressive vehicle re-identification for urban surveillance,” in ECCV . Springer, 2016, pp. 869–884
2016
Later among the works it cites.
X. Liu, W. Liu, H. Ma, and H. Fu, “Large-scale vehicle re-identification in urban surveillance videos,” in ICME , 2016, pp. 1–6
2016
Later among the works it cites.
G. Ding, Y. Guo, and J. Zhou, “Collective matrix factorization hashing for multimodal data,” in CVPR , 2014, pp. 2075–2082
2082
Closest in time.
K. Wang, R. He, W. Wang, L. Wang, and T. Tan, “Learning coupled feature spaces for cross-modal matching,” in ICCV , 2013, pp. 2088–2095
2095
Closest in time.