Fetching the paper…
Reading the bibliography…
Multi-modal data is becoming more common in big data background.
D. R. Hardoon, S. Szedmak, and J. Shawetaylor, “Canonical correlation analysis: An overview with application to learning methods,” Neural Computation , vol. 16, no. 12, pp. 2639–2664, 2004
2004
Earlier work this paper cites.
N. Rasiwasia, P. J. Moreno, and N. Vasconcelos, “Bridging the gap: Query by semantic example,” IEEE Transactions on Multimedia , vol. 9, no. 5, pp. 923–938, 2007
2007
Earlier work this paper cites.
J. Wang, S. Kumar, and S. Chang, “Sequential projection learning for hashing with compact codes,” in international conference on machine learning , 2010, Conference Proceedings
2010
Earlier work this paper cites.
C. Rashtchian, P. Young, M. Hodosh, and J. Hockenmaier, “Collecting image annotations using amazon’s mechanical turk,” in NAACL Hlt 2010 Workshop on Creating Speech and Language Data with Amazon’s Mechanical Turk , 2010, Conference Proceedings
2010
Earlier work this paper cites.
J. Ngiam, A. Khosla, M. Kim, J. Nam, H. Lee, and A. Y. Ng, “Multimodal deep learning,” in International Conference on Machine Learning, ICML 2011, Bellevue, Washington, USA, June 28 - July , 2011, Conference Proceedings
2011
Earlier work this paper cites.
Y. Jia, M. Salzmann, and T. Darrell, “Learning cross-modality similarity for multinomial data,” in International Conference on Computer Vision , 2011, Conference Proceedings
2011
Earlier work this paper cites.
S. Kumar and R. Udupa, “Learning hash functions for cross-view similarity search,” in International Joint Conference on Artificial Intelligence , 2011, Conference Proceedings
2011
Earlier work this paper cites.
A. Sharma, “Generalized multiview analysis: A discriminative latent space,” in IEEE Conference on Computer Vision and Pattern Recognition , 2012, Conference Proceedings
2012
Earlier work this paper cites.
N. Srivastava and R. Salakhutdinov, “Multimodal learning with deep boltzmann machines,” in International Conference on Neural Information Processing Systems , 2012, Conference Proceedings
2012
Earlier work this paper cites.
Z. Yi and D. Y. Yeung, “Co-regularized hashing for multimodal data,” in International Conference on Neural Information Processing Systems , 2012, Conference Proceedings
2012
Earlier work this paper cites.
K. W. Chen, N. Ayutyanont, J. B. S. Langbaum, A. S. Fleisher, C. Reschke, W. Lee, X. F. Liu, G. E. Alexander, D. Bandy, R. J. Caselli, and E. M. Reiman, “Correlations between fdg pet glucose uptake-mri gray matter volume scores and apolipoprotein e epsilon 4 gene dose in cognitively normal adults: A cross-validation study using voxel-based multi-modal partial least squares,” Neuroimage , vol. 60, no. 4, pp. 2316–2322, 2012
2012
Earlier work this paper cites.
M. Rohrbach, S. Ebert, and B. Schiele, “Transfer learning in a transductive setting,” in International Conference on Neural Information Processing Systems , 2013, Conference Proceedings
2013
Earlier work this paper cites.
G. Andrew, R. Arora, J. A. Bilmes, and K. Livescu, “Deep canonical correlation analysis,” in international conference on machine learning , 2013, Conference Proceedings
2013
Cited alongside, same era.
A. Frome, G. S. Corrado, J. Shlens, S. Bengio, J. Dean, M. Ranzato, and T. Mikolov, “Devise: A deep visual-semantic embedding model,” in neural information processing systems , 2013, Conference Proceedings
2013
Cited alongside, same era.
S. Roller and S. S. I. Walde, “A multimodal lda model integrating textual, cognitive and visual modalities,” in empirical methods in natural language processing , 2013, Conference Proceedings
2013
Cited alongside, same era.
M. Ou, P. Cui, F. Wang, J. Wang, W. Zhu, and S. Yang, “Comparing apples to oranges: a scalable solution with heterogeneous hashing,” in ACM SIGKDD International Conference on Knowledge Discovery and Data Mining , 2013, Conference Proceedings
2013
Cited alongside, same era.
Y. Guo, G. Ding, X. Jin, and J. Wang, “Transductive zero-shot recognition via shared model space learning,” in AAAI , 2016, Conference Proceedings
2016
Later among the works it cites.
Y. H. Qian, F. J. Li, J. Y. Liang, B. Liu, and C. Y. Dang, “Space structure and clustering of categorical data,” IEEE Transactions on Neural Networks and Learning Systems , vol. 27, no. 10, pp. 2047–2059, 2016
2016
Later among the works it cites.
M. N. Luo, X. J. Chang, Z. H. Li, L. Q. Nie, A. G. Hauptmann, and Q. H. Zheng, “Simple to complex cross-modal learning to rank,” Computer Vision and Image Understanding , vol. 163, pp. 67–77, 2017
2017
Later among the works it cites.
R. Socher, M. Ganjoo, C. D. Manning, and A. Y. Ng, “Zero-shot learning through cross-modal transfer,” in neural information processing systems , 2017, Conference Proceedings, pp. 935–943
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Karpathy, A. Joulin, and F. F. Li, “Deep fragment embeddings for bidirectional image sentence mapping,” in International Conference on Neural Information Processing Systems , 2014, Conference Proceedings
2014
Cited alongside, same era.
R. Socher, A. Karpathy, Q. V. Le, C. D. Manning, and A. Y. Ng, “Grounded compositional semantics for finding and describing images with sentences,” Transactions of the Association for Computational Linguistics , vol. 2, no. 0, pp. 207–218, 2014
2014
Cited alongside, same era.
Y. Wang, F. Wu, J. Song, X. Li, and Y. Zhuang, “Multi-modal mutual topic reinforce modeling for cross-media retrieval,” in acm multimedia , 2014, Conference Proceedings
2014
Cited alongside, same era.
F. Wu, Y. Zhou, Y. Yang, S. Tang, Y. Zhang, and Y. Zhuang, “Sparse multi-modal hashing,” IEEE Transactions on Multimedia , vol. 16, no. 2, pp. 427–439, 2014
2014
Cited alongside, same era.
P. J. Costa, E. Coviello, G. Doyle, N. Rasiwasia, G. R. Lanckriet, R. Levy, and N. Vasconcelos, “On the role of correlation and abstraction in cross-modal multimedia retrieval,” IEEE Transactions on Pattern Analysis & Machine Intelligence , vol. 36, no. 3, pp. 521–35, 2014
2014
Cited alongside, same era.
J. Donahue, Y. Jia, O. Vinyals, J. Hoffman, N. Zhang, E. Tzeng, and T. Darrell, “Decaf: A deep convolutional activation feature for generic visual recognition,” international conference on machine learning , pp. 647–655, 2014
2014
Cited alongside, same era.
Y. Mroueh, E. Marcheret, and V. Goel, “Multimodal retrieval with asymmetrically weighted truncated-svd canonical correlation analysis,” Computer Science , 2015
2015
Cited alongside, same era.
N. Rasiwasia, J. C. Pereira, E. Coviello, G. Doyle, G. R. G. Lanckriet, R. Levy, and N. Vasconcelos, “A new approach to cross-modal multimedia retrieval,” in International Conference on Multimedia , Conference Proceedings
Cited in the paper.
L. Wang, W. Sun, Z. Zhao, and F. Su, “Modeling intra- and inter-pair correlation via heterogeneous high-order preserving for cross-modal retrieval,” Signal Processing , vol. 131, pp. 249–260, 2017
2017
Later among the works it cites.
B. Jiang, J. Yang, Z. Lv, K. Tian, Q. Meng, and Y. Yan, “Internet cross-media retrieval based on deep learning,” Journal of Visual Communication and Image Representation , vol. 48, pp. 356–366, 2017
2017
Later among the works it cites.
Y. Wei, Y. Zhao, C. Lu, S. Wei, L. Liu, Z. Zhu, and S. Yan, “Cross-modal retrieval with cnn visual features: A new baseline,” IEEE Transactions on Systems, Man, and Cybernetics , vol. 47, pp. 449–460, 2017
2017
Later among the works it cites.
L. Huang and Y. Peng, “Cross-media retrieval by exploiting fine-grained correlation at entity level,” Neurocomputing , vol. 236, pp. 123–133, 2017
2017
Later among the works it cites.
S. Arora, Y. Liang, and T. Ma, “A simple but tough-to-beat baseline for sentence embeddings,” in International Conference on Learning Representations , 2017, Conference Proceedings
2017
Later among the works it cites.
L. Zhang, T. Xiang, and S. Gong, “Learning a deep embedding model for zero-shot learning,” computer vision and pattern recognition , pp. 3010–3019, 2017
2017
Later among the works it cites.
N. Gao, S.-J. Huang, Y. Yan, and S. Chen, “Cross modal similarity learning with active queries,” Pattern Recognition , vol. 75, pp. 214–222, 2018
2018
Closest in time.