Fetching the paper…
Reading the bibliography…
Increasingly many real world tasks involve data in multiple modalities or views.
Relations between two sets of variates
H. Hotelling · 1936
Earlier work this paper cites.
A general class of coefficients of divergence of one distribution from another
S. M. Ali and S. D. Silvey · 1966
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Wordseye: An automatic text-to-scene conversion system
B. Coyne and R. Sproat · 2001
Earlier work this paper cites.
On kernel-target alignment
N. Cristianini, J. Shawe-Taylor, A. Elisseeff, and J. S. Kandola · 2002
Earlier work this paper cites.
Kernel independent component analysis
F. R. Bach and M. I. Jordan · 2003
Earlier work this paper cites.
Canonical correlation analysis: An overview with application to learning methods
D. R. Hardoon, S. R. Szedmak, and J. R. Shawe-taylor · 2004
Earlier work this paper cites.
Kernelizing sorting, permutation and alignment for minimum volume PCA
T. Jebara · 2004
Earlier work this paper cites.
Measuring statistical dependence with Hilbert-schmidt norms
A. Gretton, O. Bousquet, A. Smola, and B. Schölkopf · 2005
Earlier work this paper cites.
Learning bilingual lexicons from monolingual corpora
A. Haghighi, P. Liang, T. Berg-Kirkpatrick, and D. Klein · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
Learning to detect unseen object classes by between-class attribute transfer
C. H. Lampert, H. Nickisch, and S. Harmeling · 2009
Earlier work this paper cites.
Kernelized sorting
N. Quadrianto, L. Song, and A. J. Smola · 2009
Earlier work this paper cites.
Kernelized sorting for natural language processing
J. Jagarlamudi, S. Juarez, and H. Daumé III · 2010
Earlier work this paper cites.
Sufficient dimension reduction via squared-loss mutual information estimation
T. Suzuki and M. Sugiyama · 2010
Earlier work this paper cites.
The importance of encoding versus training with sparse coding and vector quantization
A. Coates and A. Y. Ng · 2011
Earlier work this paper cites.
Multimodal deep learning
J. Ngiam, A. Khosla, M. Kim, J. Nam, H. Lee, and A. Y. Ng · 2011
Cited alongside, same era.
Im2text: Describing images using 1 million captioned photographs
V. Ordonez, G. Kulkarni, and T. L. Berg · 2011
Cited alongside, same era.
Cross-domain object matching with model selection
M. Yamada and M. Sugiyama · 2011
Cited alongside, same era.
Convex kernelized sorting
N. Djuric, M. Grbovic, and S. Vucetic · 2012
Cited alongside, same era.
Choosing linguistics over vision to describe images
A. Gupta, Y. Verma, and C. V. Jawahar · 2012
Cited alongside, same era.
Deep canonical correlation analysis
G. Andrew, R. Arora, J. Bilmes, and K. Livescu · 2013
Cited alongside, same era.
Attribute-based classification for zero-shot visual object categorization
C. H. Lampert, H. Nickisch, and S. Harmeling · 2014
Later among the works it cites.
Microsoft COCO: Common Objects in Context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
P. Young, A. Lai, M. Hodosh, and J. Hockenmaier · 2014
Later among the works it cites.
Transductive multi-view zero-shot learning
Y. Fu, T. Hospedales, T. Xiang, and S. Gong · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
B. Chang, U. Kruger, R. Kustra, and J. Zhang · 2013
Cited alongside, same era.
Devise: A deep visual-semantic embedding model
A. Frome, G. Corrado, J. Shlens, S. Bengio, J. Dean, M. Ranzato, and T. Mikolov · 2013
Cited alongside, same era.
A multi-view embedding space for modeling internet images, tags, and their semantics
Y. Gong, Q. Ke, M. Isard, and S. Lazebnik · 2013
Cited alongside, same era.
YouTube2Text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition
S. Guadarrama, N. Krishnamoorthy, G. Malkarnenkar, S. Venugopalan, R. Mooney, T. Darrell, and K. Saenko · 2013
Cited alongside, same era.
Unsupervised cluster matching via probabilistic latent variable models
T. Iwata, T. Hirao, and N. Ueda · 2013
Cited alongside, same era.
Generating natural-language video descriptions using text-mined knowledge
N. Krishnamoorthy, G. Malkarnenkar, R. Mooney, K. Saenko, and S. Guadarrama · 2013
Cited alongside, same era.
B. Klein, G. Lev, G. Sadeh, and L. Wolf · 2015
Later among the works it cites.
Cross-domain matching with squared-loss mutual information
M. Yamada, L. Sigal, M. Raptis, M. Toyoda, Y. Chang, and M. Sugiyama · 2015
Later among the works it cites.
Deep correlation for matching images and text
F. Yan and K. Mikolajczyk · 2015
Later among the works it cites.
Correlational neural networks
S. Chandar, M. M. Khapra, H. Larochelle, and B. Ravindran · 2016
Later among the works it cites.
Multi-view deep network for cross-view classification
M. Kan, S. Shan, and X. Chen · 2016
Later among the works it cites.
A survey on heterogeneous face recognition: Sketch, infra-red, 3d and low-resolution
S. Ouyang, T. Hospedales, Y.-Z. Song, X. Li, C. C. Loy, and X. Wang · 2016
Later among the works it cites.
Learning deep structure-preserving image-text embeddings
L. Wang, Y. Li, and S. Lazebnik · 2016
Later among the works it cites.
Word translation without parallel data
A. Conneau, G. Lample, M. Ranzato, L. Denoyer, and H. Jégou · 2017
Closest in time.
Deep visual-semantic alignments for generating image descriptions
A. Karpathy and L. Fei-Fei · 2017
Closest in time.
Learning robust visual-semantic embeddings
Y. H. Tsai, L. Huang, and R. Salakhutdinov · 2017
Closest in time.
Learning two-branch neural networks for image-text matching tasks
L. Wang, Y. Li, and S. Lazebnik · 2017
Closest in time.